← Latest editionRSS

Independent daily intelligence

Signal Archive

47 editions on record

  1. 047
    StoreBench: a live-commerce environment for autonomous operator agents
    24 stories →
  2. 046
    StoreBench: a live-commerce environment for autonomous operator agents
    19 stories →
  3. 045
    World Labs launches a public API for generating explorable 3D worlds
    22 stories →
  4. 044
    OpenTPU: an open-source AI accelerator
    20 stories →
  5. 043
    Beam: Reflection's 501B open-weight model
    24 stories →
  6. 042
    Seeing, Saying, but Not Using: From Reportable Spatial Facts to Usable States in Multimodal Large Language Models
    21 stories →
  7. 041
    Kolibri: A Sovereign Open-Weight Model
    19 stories →
  8. 040
    Announcing the World API
    22 stories →
  9. 039
    Clef: Open-weight decision models and a new RL fine-tuning platform
    22 stories →
  10. 038
    ArgGYM: A Procedural, Engine-Verified Benchmark for Structured Defeasible Reasoning
    23 stories →
  11. 037
    PSSA: A non-transformer language model written from scratch in Rust
    24 stories →
  12. 036
    Streaming 3DGS worlds on the web
    24 stories →
  13. 035
    The Price of Thought: Does Test-Time Reasoning Pay in LLM Trading?
    23 stories →
  14. 034
    Reverse-engineering the vintage Intel 8087 tangent algorithm: more than CORDIC
    23 stories →
  15. 033
    Thinking Leakage: A causal audit of NoThink post-training in hybrid reasoning models
    14 stories →
  16. 032
    AgenticCADedit: Stateful, tool-mediated multimodal CAD editing
    24 stories →
  17. 031
    COPE: continual LLM personalization from sparse feedback
    19 stories →
  18. 030
    The economics of open-weight inference
    21 stories →
  19. 029
    GVPO++: Group Variance Policy Optimization for LLM Post-Training and On-Policy Distillation
    18 stories →
  20. 028
    GVPO++: Group Variance Policy Optimization for LLM post-training and on-policy distillation
    20 stories →
  21. 027
    Saving another 100TB of RAM with math and Rust
    23 stories →
  22. 026
    Saving another 100TB of RAM with math and Rust
    23 stories →
  23. 025
    Training a 4B model to produce 81% faster query plans than Postgres
    24 stories →
  24. 024
    Announcing the World API
    22 stories →
  25. 023
    Distributed Systems Classics (2017)
    16 stories →
  26. 022
    Mapping the mind of a large language model
    23 stories →
  27. 021
    Anthropic maps concepts inside Claude Sonnet
    16 stories →
  28. 020
    Mapping the mind of a large language model
    24 stories →
  29. 019
    T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks
    24 stories →
  30. 018
    Streaming 3DGS worlds on the web
    17 stories →
  31. 017
    When memory helps tool-using LLM agents
    24 stories →
  32. 016
    Inside OpenAI’s research acceleration with coding agents
    22 stories →
  33. 015
    Extremely Sparse Supervision Incentivizes Reasoning Ability
    24 stories →
  34. 014
    Streaming 3DGS worlds on the web
    24 stories →
  35. 013
    World Labs launches a public World API
    17 stories →
  36. 012
    Speculative macro commit for faster tool-using agents
    18 stories →
  37. 011
    The efficient frontier of LLM inference
    18 stories →
  38. 010
    GPU World
    23 stories →
  39. 009
    How to build a diffusion language model
    20 stories →
  40. 008
    How to build a diffusion language model
    23 stories →
  41. 007
    GLM-5.3 is now open-weight
    21 stories →
  42. 006
    GLM-5.3 is now open-weight
    24 stories →
  43. 005
    GRAS: Guided Reduced-Variance Proposals and Adaptive Selection for Training-Free Reward Alignment in Discrete Diffusion
    24 stories →
  44. 004
    SHIFT-LLM: correcting distribution shift after depth pruning
    22 stories →
  45. 003
    Evaluating LLMs as calibrated causal-edge classifiers
    23 stories →
  46. 002
    Executable Is a SQLite Database
    23 stories →
  47. 001
    Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things
    23 stories →