AI
Breaking AI announcements, model launches, company news
The lead
Latest firstBrowse AI by topic
Open-Source AI
Hugging Face, GitHub, Ollama, local models, agent frameworks
Explore 55 postsModel Launches
Every AI model release that matters: benchmarks, context windows, pricing, and honest first takes.
Explore 43 postsAI Agents & MCP
Autonomous AI agent frameworks, multi-agent systems, and real deployments shipping today.
Explore 29 postsAI Use Cases
Real examples of what people are doing with AI
Explore 22 postsResearch
Peer-reviewed AI research distilled: what was proven, what it changes, and what it does not.
Explore 10 postsPolicy & Safety
AI legislation, safety frameworks, and court cases shaping how AI gets deployed.
Explore 1 postAI Tools
Curated directory of AI tools by category
Explore 1 postLocal AI
Local AI installs and private inference: Ollama, llama.cpp, and self-hosted setups.
ExploreLatest in AI
New metric reveals AI agents behave inconsistently
A new arXiv metric shows AI agents can pass tasks yet behave unpredictably across them, exposing a reliability gap that success-rate scores…
AI coding ability jumped 6x a year as costs fell
An arXiv review follows AI from BERT in 2018 to frontier agents in 2026, finding coding ability up sixfold a year while…
One INT8 kernel swap changes every vLLM output
A new arXiv preprint swaps CUTLASS for Triton inside vLLM and finds no matching sequences, tracing the drift to scale rounding, not…
Agentao lets hosts veto every AI agent action before it runs
Agentao proposes a local-first runtime that splits an LLM's action proposals from host-authorized execution, adding permissions, replay, and audit trails.
arXiv paper scores AI harnesses without labels
New arXiv framework judges if an LLM-with-memory agent improves by measuring how fast a student converges to a stronger teacher, no labeled…
New Transformer swaps global attention for block memory
An arXiv paper shows BCMT matching dense Transformer quality on long context while raising training throughput and lowering memory use.
Batch pruning keeps reasoning models fast at scale
A training-free pruning method lets reasoning models keep accuracy and speed under batched inference, beating prior art by 39.7 points at 50%…
Microsoft patches 10 .NET flaws in August 2026 update
Microsoft's August 2026 .NET servicing update fixes 10 CVEs across .NET 8, 9, and 10, including remote code execution and privilege flaws.…
arXiv agent-memory model bars stale, retracted data
Researchers propose Governed Persistent Memory, an auditable model that blocks agents from releasing claims built on retracted, deleted, or stale records.
OpenAI opens its cyber models to 16 security firms
Accenture, IBM, Cloudflare and 13 other partners can now run OpenAI's controlled-access cyber models inside the security services their clients already buy.