Laya runs offline on M4, ChatGPT tracks ads
New offline inference benchmarks, cross‑site data collection, and rescue services raise fresh questions for AI developers.
New offline inference benchmarks, cross‑site data collection, and rescue services raise fresh questions for AI developers.
Cognition launches SWE‑2, DeepSeek releases v4.1 Flash, and a community project trains a 3.8B model for under $1k, reshaping cost expectations.
A 27B Qwen model was reverse‑engineered in half an hour, while a NanoGPT speedrun on Frontier pushes training speed. The piece examines why local LLMs still lag behind cloud services.
A rundown of recent Hacker News threads on a Claude feature request, Unsloth's GGUF release, Opus 5.0's incoherence issue, and a blog on extensible software for LLMs.
OpenAI's o3 model scores 88% on ARC-AGI, prompting a reckoning with static benchmarks that are rapidly losing relevance for measuring LLM progress.
Manifest shuts down its router while open‑source alternatives like RouteLLM, Any‑LLM, and ClawRouter vie for the cost‑saving niche.
OpenAI's desktop app falters, Claude suffers errors, and copyright suits against generative AI rise, while moderation workarounds surface.
A blog shows the open‑source Gemma‑4 LLM operating on a 2016 Xeon, sparking debate over hardware versus software in AI deployment.
Local AI chatbots now run on iPhone, ElevenLabs adds Stan Lee’s voice, and new research shows frontier LLMs often disagree on factual claims.
New Rust‑based clients and LLM retrospectives deepen engagement with Hacker News while critics spotlight tech labor tactics.
Antigravity 2.0 tops an OpenSCAD LLM test, Wozniak backs AI intelligence, Slumber adds a TUI HTTP client, and Cleve Moler passes away.
OpenAI prioritizes coding & enterprise, former Sora boss leaves