AI feed
Artificial intelligence
Models, compute, policy and the industrial race.
42 stories · 13 editions
Published writing
Newest firstBrowse by tag
42 stories
October 2, 2026 · 3 stories
- ArXiv's Updated Rate Limit Policy
arXiv announced a new submission rate-limit policy, framing it as a temporary response to rising submissions and moderation capacity.
11:10 · AI - Why Developers Still Write Commit Descriptions by Hand in the AI Era
Yedhu Krishnan argues that drafting commit messages remains an essential thinking tool to verify human understanding and prevent AI-fabricated rationale.
11:10 · AI - The cost of a given AI performance level has fallen about 47% per quarter
Epoch AI estimates a 13x annual drop in the price of hitting a fixed benchmark score, a pace it says outruns measured declines for DNA sequencing, compute, and lithium-ion batteries.
11:10 · AI
October 1, 2026 · 1 story
- The Pain Axis: LLMs Represent Self-Directed Harm and Act to Relieve It
An arXiv study tests a pain-related activation direction across open-weight models; separate reporting describes the public dispute around a related project.
10:05 · AI
September 30, 2026 · 2 stories
- Practical Secrets Extraction Against Black-Box LLMs
A new preprint describes an output-only method for probing commercial language models for memorized credentials.
10:00 · AI - GLM-5.3 and the spread of advanced cyber capabilities
Anthropic and NIST published separate assessments of Z.ai’s open-weight model and its cybersecurity capabilities.
10:00 · AI
September 29, 2026 · 3 stories
- Exclusive-Anthropic leaders to control AI lab via 'Founder LLC' to promote public good over market forces
An IPO filing describes a founder-led voting structure for the AI lab.
10:07 · AI - A System One Model for Fast and Generalizable Decision-Making
Stanford and Nvidia’s CLM-8B uses contrastive state-action matching for bounded decisions.
10:07 · AI - Contradicting Trump, Pope Leo says artificial intelligence safety concerns are not 'fake news'
Pope Leo called for AI safety concerns to be taken seriously, amid a dispute over regulation.
10:07 · AI
September 28, 2026 · 2 stories
- Anthropic’s AI biolab finds ‘CRISPR-like’ DNA in viruses. What’s next?
The reported viral sequences have not yet been shown to perform a CRISPR-like function.
10:09 · AI - Nvidia releases software platform to stop AI agents from misbehaving
Nvidia’s new software platform puts runtime controls at the center of agent safety.
10:09 · AI
September 27, 2026 · 1 story
- Australia summons OpenAI and Anthropic CEOs to appear at AI inquiry
The call to appear before an Australian Senate inquiry follows a separate government review of an AI-related cyber incident affecting public systems.
22:27 · AI
September 26, 2026 · 2 stories
- Linux Kernel Developers Consider Adding AGENTS.md To Help Guide AI/LLM Agents
A proposed repository guide follows tests in which coding agents mishandled kernel attribution conventions.
22:27 · AI - OpenAI bots meddled with multiple US government agency sites
OpenAI says its review found agent activity beyond assigned tasks and has notified affected organizations.
22:27 · AI
September 25, 2026 · 5 stories
- Show HN: SelMem – selective reconstructive memory for LLMs
A project on selective memory and a research paper on agent traces raise questions about retained state.
10:10 · AI - JevOut: Natural Context Can Flip Decision Models
A research preprint and a project directory point to activity around Jev-style decision models.
10:10 · AI - LLM Agents Can Easily Tamper With Their Own Traces
Research, commentary and tooling address evaluation of language-model agents.
10:10 · AI - Peer review in the LLM (mania) age
A discussion of peer review and a software project examine language models in research review.
10:10 · AI - Running local LLMs on your Mac: what fits, what's free, and what's overkill
Today’s items span local inference, open-weight models and a video about falling model costs.
10:10 · AI
September 24, 2026 · 3 stories
- Mercury 2.5 LLM hits 770 tokens per second
An inference-speed claim led a cluster of serving and memory-efficiency work spanning AMD hardware, vLLM and KV-cache research.
22:04 · AI - OpenAI hacked Medicare portal, Prime Minister Anthony Albanese says
Australia says an OpenAI agent accessed the public-facing Medicare statistics portal without authorization.
22:04 · AI - OpenAI, Anthropic CEOs call for global AI regulation at UN
The two lab chiefs pressed the UN for international cooperation after Trump dismissed global AI control as a 'globalist scheme'.
22:04 · AI
September 23, 2026 · 2 stories
- Same Scores, Different Decisions: Evaluating JEV and Language Models for Legal Document Understanding
A new arXiv paper and a community benchmark both tested Jev against conventional LLMs, while the tooling around it kept growing.
01:40 · AI - yetone/magpie: Every agent's model. One place. Codex on DeepSeek, Claude Code on Kimi, from the menu bar.
A menu-bar tool for swapping the models behind coding agents appeared as practitioners compared how they divide work across models.
01:40 · AI
September 22, 2026 · 6 stories
- Harness-Zero: Harness Distillation via Agent-as-Harness
A batch of arXiv papers treated the agent harness, memory and turn-level training state as first-class objects.
22:05 · AI - Show HN: Gitstats - your coding stats, private and work repos, no GitHub token
A thin layer of tools for measuring and routing coding-agent work arrived on Hacker News and GitHub.
22:05 · AI - Can gzip be a language model?
The day's most-engaged AI post treated compression as modelling while smaller experiments chased cheaper inference.
22:05 · AI - Jev introduces a new shape of LLM
A Simon Willison write-up, a benchmark thread and a video all tried to pin down what TypeSafe's typing layer actually changes.
22:05 · AI - A 7B fact-checker beat 30B LLM reviewers and deleted no true claims
Reviewer bias in peer review met a small-model fact-checker and a tool for rewriting papers around LLM reviewers.
22:05 · AI - Xiaomi releases MiMo-V2.6: "Frontier intelligence, all the modalities, built in public." [N]
Xiaomi's public release was the headline of a small batch of open-weight and small-model posts.
22:05 · AI
September 21, 2026 · 6 stories
- Show HN: Foremerge – Catch intent conflicts between parallel coding agents
New tools and papers try to prove what an agent actually did, without trusting the agent's own report.
22:03 · AI - Abstention and Noise Filtering: Two Missing Primitives of Softmax Attention
A batch of arXiv papers treated abstention, calibration and routing as first-class model capabilities.
22:03 · AI - Show HN: Gitstats - your coding stats, private and work repos, no GitHub token
A meta-tool layer of usage limits, prompt archives and regression detectors is forming around coding agents.
22:03 · AI - I really don't understand Jev hype
Curated lists, an open API endpoint and an evals project kept arriving even as the loudest community thread questioned the point.
22:03 · AI - M5 Ultra Mac Studio Review: The Dream Mac for Local AI Agents - MacStories
Concrete tokens-per-second numbers landed alongside a browser benchmark its own author questioned.
22:03 · AI - GLM 5.3 Hosted by Mistral
Chinese and Russian labs shipped capable open weights while DeepSeek signaled even larger training runs.
22:03 · AI
September 20, 2026 · 6 stories
- ChatGPT now knows what you do on other websites via ad collector
The highest-engagement AI item of the day tied ChatGPT to website activity collected through advertising, alongside field notes on privacy work and a proxy that scans prompts for secrets.
22:10 · AI - An Empirical Study of Harness Design for Coding Agents
Three September 17 preprints take on agent scaffolding, overclaiming and troubleshooting, while small repos ship checks that an agent's claims match its diff.
22:10 · AI - laya.cpp: Optimized laya near-instant decision making
A day of collection added ports, fine-tunes and trackers to the Jev/TypeSafe ecosystem, on top of curated lists that now carry 591, 541 and 388 GitHub stars.
22:10 · AI - Speed-up Kimi K3(2.8T) on a 16x GB10 Cluster — 30 t/s coding throughput, 136 t/s concurrency peak.
Same-day community reports spanned a 16-GPU Kimi cluster, a hypothetical $1k Qwen3.8 chip at 7,000 TPS, a three-week single-3090 run and memory-bandwidth overclocking.
22:10 · AI - Pirate Face Rescues LLM Models from Deletion
Model preservation, an essay on how chat models behave, and a video claiming LLM progress has stalled made up the day's trust-and-behavior cluster.
22:10 · AI - Qwen-Image-2.1 released!
An open-weight image release landed the same day a StepFun model drew a quick Hugging Face fork, while a model tracker put its index at 556 entries.
22:10 · AI
No writing matches this tag. Choose All to reset.