Show HN:我们对18个大型语言模型进行了OCR基准测试(7000多次调用)——更便宜的模型胜出
Arbitr HQ对18个主流大型语言模型(LLMs)进行了超过7000次调用的OCR(光学字符识别)基准测试。结果显示,成本更低的模型在OCR任务中表现优于高价模型,为开发者在选择LLM进行文字识别时提供了性价比参考。
Arbitr HQ对18个主流大型语言模型(LLMs)进行了超过7000次调用的OCR(光学字符识别)基准测试。结果显示,成本更低的模型在OCR任务中表现优于高价模型,为开发者在选择LLM进行文字识别时提供了性价比参考。
A new phishing-as-a-service called Starkiller uses disguised links to load real login pages from target brands. It acts as a relay between victims and legitimate sites, forwarding usernames, passwords, and MFA codes to bypass security measures.
An investigation uncovered a large network of fake support groups on Telegram that spread cryptocurrency stealers and drainers. The network was found to be actively promoting malicious tools designed to drain crypto wallets.
Gemini can identify public figures in images, while ChatGPT and Claude currently do not offer this capability. This represents a functional difference between major AI models regarding image recognition of people.
Inception Labs has launched Mercury 2, described as the world's first reasoning diffusion LLM. The diffusion language model reportedly delivers 5x faster inference speed compared to leading speed-optimized LLMs.
Andrej Karpathy describes using LLMs to build personal knowledge bases by indexing source documents into a raw directory, then having the LLM compile them into a markdown wiki with summaries, backlinks, and categorization. The system allows for complex Q&A against the wiki and can generate various output formats like markdown files, slideshows, and images, all viewable in Obsidian.