Skip to content

AI feed

Artificial intelligence

Models, compute, policy and the industrial race.

18 stories 107 source citations 3 editions Default view: dense list Toggle any time

Lead

Widest cluster coverage in this feed
AI DEEP September 22, 2026 10:46

Jev introduces a new shape of LLM

A Simon Willison write-up, a benchmark thread and a video all tried to pin down what TypeSafe's typing layer actually changes.

TL;DR

  1. Simon Willison's write-up said Jev introduces a new shape of LLM, treating TypeSafe's System One as a typed decision layer distinct from a chat model.
  2. An r/MachineLearning thread reported that Jev's calibration was measured and the LLMs won, while a Two Minute Papers episode claimed a 200x speed-up with a caveat.
  3. An arXiv paper evaluated Jev for scientific decisions, and an ecosystem submission argued any LLM can be used just like JEV.
4 sources · 5 min · cluster: 3 · jev typesafe calibration

AI timeline

Newest first · UTC
View

September 22, 2026 · 5 stories

  1. Can gzip be a language model?

    The day's most-engaged AI post treated compression as modelling while smaller experiments chased cheaper inference.

    AI DEEP 4 sources · 4 min · cluster 3
    10:46 · AI
  2. Harness-Zero: Harness Distillation via Agent-as-Harness

    A batch of arXiv papers treated the agent harness, memory and turn-level training state as first-class objects.

    AI DATA 2 sources · 3 min · cluster 2
    10:46 · AI
  3. A 7B fact-checker beat 30B LLM reviewers and deleted no true claims

    Reviewer bias in peer review met a small-model fact-checker and a tool for rewriting papers around LLM reviewers.

    AI DEEP 3 sources · 4 min · cluster 2
    10:46 · AI
  4. Xiaomi releases MiMo-V2.6: "Frontier intelligence, all the modalities, built in public." [N]

    Xiaomi's public release was the headline of a small batch of open-weight and small-model posts.

    AI BRIEF 2 sources · 3 min · cluster 2
    10:46 · AI
  5. Show HN: Gitstats - your coding stats, private and work repos, no GitHub token

    A thin layer of tools for measuring and routing coding-agent work arrived on Hacker News and GitHub.

    AI BRIEF 3 sources · 3 min · cluster 3
    10:46 · AI

September 21, 2026 · 6 stories

  1. I really don't understand Jev hype

    Curated lists, an open API endpoint and an evals project kept arriving even as the loudest community thread questioned the point.

    AI DEEP 3 sources · 5 min · cluster 3
    22:03 · AI
  2. GLM 5.3 Hosted by Mistral

    Chinese and Russian labs shipped capable open weights while DeepSeek signaled even larger training runs.

    AI DEEP 3 sources · 4 min · cluster 3
    22:03 · AI
  3. M5 Ultra Mac Studio Review: The Dream Mac for Local AI Agents - MacStories

    Concrete tokens-per-second numbers landed alongside a browser benchmark its own author questioned.

    AI DEEP 4 sources · 4 min · cluster 3
    22:03 · AI
  4. Show HN: Foremerge – Catch intent conflicts between parallel coding agents

    New tools and papers try to prove what an agent actually did, without trusting the agent's own report.

    AI DEEP 2 sources · 4 min · cluster 3
    22:03 · AI
  5. Abstention and Noise Filtering: Two Missing Primitives of Softmax Attention

    A batch of arXiv papers treated abstention, calibration and routing as first-class model capabilities.

    AI DATA 2 sources · 3 min · cluster 2
    22:03 · AI
  6. Show HN: Gitstats - your coding stats, private and work repos, no GitHub token

    A meta-tool layer of usage limits, prompt archives and regression detectors is forming around coding agents.

    AI BRIEF 3 sources · 3 min · cluster 2
    22:03 · AI

September 20, 2026 · 6 stories

  1. laya.cpp: Optimized laya near-instant decision making

    A day of collection added ports, fine-tunes and trackers to the Jev/TypeSafe ecosystem, on top of curated lists that now carry 591, 541 and 388 GitHub stars.

    AI DEEP 3 sources · 5 min · cluster 3
    22:10 · AI
  2. ChatGPT now knows what you do on other websites via ad collector

    The highest-engagement AI item of the day tied ChatGPT to website activity collected through advertising, alongside field notes on privacy work and a proxy that scans prompts for secrets.

    AI BRIEF 3 sources · 3 min · cluster 2
    22:10 · AI
  3. Qwen-Image-2.1 released!

    An open-weight image release landed the same day a StepFun model drew a quick Hugging Face fork, while a model tracker put its index at 556 entries.

    AI BRIEF 2 sources · 3 min · cluster 2
    22:10 · AI
  4. Speed-up Kimi K3(2.8T) on a 16x GB10 Cluster — 30 t/s coding throughput, 136 t/s concurrency peak.

    Same-day community reports spanned a 16-GPU Kimi cluster, a hypothetical $1k Qwen3.8 chip at 7,000 TPS, a three-week single-3090 run and memory-bandwidth overclocking.

    AI DATA 2 sources · 4 min · cluster 2
    22:10 · AI
  5. An Empirical Study of Harness Design for Coding Agents

    Three September 17 preprints take on agent scaffolding, overclaiming and troubleshooting, while small repos ship checks that an agent's claims match its diff.

    AI DEEP 3 sources · 5 min · cluster 2
    22:10 · AI
  6. Pirate Face Rescues LLM Models from Deletion

    Model preservation, an essay on how chat models behave, and a video claiming LLM progress has stalled made up the day's trust-and-behavior cluster.

    AI BRIEF 3 sources · 4 min · cluster 2
    22:10 · AI

Type to search

↑↓ navigate ↵ open esc close