Skip to content

#Open-source models

0 items today
4/13Mon
4/4Sat
  1. Hacker News · Prompt Injection52

    PIGuard:通过 MOF 策略缓解提示词注入防护的过度防御

    圣路易斯华盛顿大学与威斯康星大学麦迪逊分校的研究者提出 PIGuard,一个用于检测提示词注入的轻量防护模型,并配套发布 NotInject 评测数据集。NotInject 包含 339 条带触发词的良性样本,用于衡量防护模型的过度防御问题,结果显示现有 SOTA 模型准确率降至接近随机猜测的 60%。

    Awaiting translation

3/17Tue
  1. Geoffrey Huntley · Blog31

    AI 作为经济战:开源模型如何成为国家间的金融武器

    Geoffrey Huntley 认为开源模型正被用作国家间的经济武器:中国厂商免费或以约 5 美元/月开放前沿开源模型,而美国向头部实验室投入数万亿美元,本地模型目前仅落后前沿实验室两到四个月。他建议预算有限的学生使用这些开源模型,企业则应按本地推理可用的思路构建业务,并提醒信任与断供风险同样适用于前沿实验室。

    Awaiting translation

3/13Fri
  1. Martin Alderson78

    How to OCR Documents with Qwen 3.5 Series Models

    The author used the open-source multimodal Qwen 3.5 series models for PDF OCR: first exporting each page as an image at 100 dpi with PyMuPDF, then feeding the images to the model for recognition. In testing, Qwen3.5-9B hit the sweet spot between quality and speed, while the smaller 0.8B to 2B models tended to go off track on complex documents, summarizing the content instead of transcribing it.

    Why it matters: The author tested Qwen 3.5 models of various sizes for PDF OCR, and shares two reusable paths—local and via OpenRouter—along with cost data.

2/15Sun