Anthropic launched its Cyber Mission, including OSS Scanner, a free opt-in service that scans open-source projects with frontier Claude models and returns a proof-of-concept, an explanation and a suggested fix. Read: Anthropic launched its Cyber Mission, including OSS Scanner, a free opt-in service that scans open-source projects with frontier Claude models and returns a proof-of-concept, an explanation and a suggested fix. Read: Anthropic cut Sonnet 5.5 cache-read pricing on the Claude Platform to $0.10 per million tokens, half the previous rate. Input stays at $2 and output at $10, and Claude Code usage limits do not change. Try: OpenAI's math repository withdrew three algebraic geometry manuscripts after a sign error broke a construction they shared, repaired proofs across other papers, and now has about 42% of headline results formalized in Lean. Read: Inception updated its Mercury Decide decision model with better benchmark results, context doubled from 33k to 66k tokens, and general availability with zero data retention at $0.02 per million input tokens. Read: OpenAI is rolling out Ultrafast mode for GPT-6.1 Sol across the API, Codex and ChatGPT Work, running up to 8x faster at $12 input and $60 output per million tokens, with US and EU data residency supported. Read: Artificial Analysis and Harvey changed the Legal Agent Benchmark to credit a task only when it meets every rubric item with no material hallucination. Grok 4.7 now leads at 9.4%, and most passing runs fail the gate. Read: A new empirical study measured agent Skills against a no-Skill baseline across nine model and harness configurations. Skills flipped from helpful to harmful on more than a third of tasks, depending on the setup.
Open-source maintainers can now get free Claude vulnerability scans
Anthropic launched its Cyber Mission, including OSS Scanner, a free opt-in service that scans open-source projects with frontier Claude models and returns a proof-of-concept, an explanation and a suggested fix. Anthropic also cut Sonnet 5.5 cache-read pricing on the Claude Platform to $0.10 per million tokens, half the previous rate, while input stays at $2 and output at $10 and Claude Code usage limits do not change. A new empirical study measured agent Skills against a no-Skill baseline across nine model and harness configurations and found Skills flipped from helpful to harmful on more than a third of tasks, depending on the setup.
Run your AI side-project on zahid.host
EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.