AI NEWS DAILY

Hardware

Hands-On with the AMD Ryzen AI Halo
Ultra budget 20GB vram with 448GB/s for $100 bucks.
Benchmark - 4x 5060 Ti (64GB VRAM) (P2P) - Qwen3.6 27B (INT8 /w bf16 kv…
**Your $80 Tesla P100 has been doing silently noisy math in llama.cpp…

Breaking | Apple sues OpenAI, accuses ex-employees of stealing trade secrets
Breaking | GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture
Old and new apps, via modern coding agents by Terry Tao
Mesh LLM: distributed AI computing on iroh
Ghost Font: A font that humans can read but AI cannot
Stop Telling Me to Ask an LLM
How the terrorist group Boko Haram uses frontier AI
AI Boosts Research Careers but Flattens Scientific Discovery
GPT-5.6, Grok 4.5, Claude, and Muse Spark build the same 4 apps
Reverse centaurs are the answer to the AI paradox (2025)
Who manages the agents?
Apple sues OpenAI, accusing it of stealing company secrets
Wealthy AI workers send San Francisco house prices soaring
AI Can't Recreate the Thrust Game (But It Can Help You Understand It)
Microsoft latest report shows 25% emissions raised due to AI data…
Meta pulls new AI image feature after days of backlash
Companies are scrambling to curtail soaring AI costs
Ask HN: How do you use Vim in the era of AI?
GPT-5.6
OpenAI released GPT-5.6, their latest model iteration with significant capability improvements across reasoning, coding, and multimodal tasks.
AI 2040: Plan A
AI Futures Project publishes 'Plan A' — a scenario for delaying superintelligence until 2040 through international cooperation, total research transparency, and mutually assured compute destruction.
AI-generated videos to maximally drive a target brain region
EPFL researchers developed AI-generated videos designed to maximally activate a target brain region, using fMRI data to optimize visual stimuli.
ChatGPT Work
OpenAI launched ChatGPT Work, an enterprise version of ChatGPT designed for professional workflows and organizational use.
AI content is everywhere on social media, especially LinkedIn
Pangram study finds AI-generated content is pervasive across social media, with LinkedIn as the worst offender.
Building a real-time AI tutor for 5-year-olds
Ello blog post details building a real-time AI tutor optimized for 5-year-old learners with 1000ms latency.
Ben Bernanke Joins Anthropic Oversight Trust
SimPolitics: America’s quest to solve politics with computers
Show HN: Reverse-engineering web apps into agent tools
How version control will evolve for the agent boom
Show HN: Reviving my 2001 college band with AI
The next era of AI is about infrastructure, not just models
I think I have LLM burnout
Developer reports chronic 'LLM burnout' from hours of daily interaction with AI assistants, describing a shift from writing code to designing, prompting, and reviewing AI-generated code.
SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
Cognition launched SWE-1.7, reaching frontier-level intelligence (near GPT-5.5 and Opus) at much lower cost, trained from a Kimi K2.7 base with extensive RL post-training.
Suspecting AI cheating, Ivy League prof ordered in-person final; scores…
Brown University professor suspected AI cheating on take-home exams; in-person final scores dropped 50%.
We made Grok 4.5, GPT-5.5, and Claude build the same apps
Side-by-side comparison of Grok 4.5, GPT-5.5, and Claude building identical applications.
What's slowing down the AI buildout
Benchmarking coding agents on Databricks' multi-million line codebase
Databricks benchmarks AI coding agents against their own multi-million-line production codebase.
AI changes the economics of software rewrites
Ask HN: Another "Hacker News" with less AI and more human-focused…
Hacker News discussion seeking alternative tech news platforms with less AI content and more human hacking.
I didn't give up - extGemma4-40_5B returned
Zer0Fit: I took Google's new TabFM & TimesFM ML foundation models and…
I benched quad 5060Tis for code generation with Qwen3.6-27B so you…
Performance comparison on full compute performance (Anima) and LLM…
OpenCoF: Learning to Reason Through Video Generation
Voodoo Quant beats Unsloth Dynamic 2.0 KLD by 95% in Qwen3.5 0.8B and 2B
If you use Open Code or other agenting programs you are leaving a lot…
I mapped Anthropic’s J-Space Hallucination signal across 7 datasets on…
Need help tuning cache in llama-server
First attempts at a CPU setup - MS-02 Intel 285hx, trying Qwen3…
LinkedIn is the undisputed king of long-form AI slop, according to a…
Claude Code now has a built-in browser that lets the AI read, click…
ARDY: Autoregressive Diffusion with Hybrid Representation for…
S&P Global sees OpenAI as a "key credit risk" for Oracle and cuts its…
Meta kills Muse Image feature that let anyone generate AI photos of…
OpenAI CEO Altman is now "pretty sure" AI is net job-creating, which is…
Interactive Jacobian-Lens visualizer and live steerer for GGUF models…
i would like to share my experience. working with huge LLMs and poor…
Working around Qwen3.6-27B's tool-call failures and looping
Kreuzberg (local document extraction) is being renamed to Xberg…
Qwenthropic
Claude Cowork's biggest use case is the mundane office work nobody…
AI agents win at Slay the Spire 2 after researchers replace growing…
Grades dropped from 96 to 48 percent when a Brown professor made…
GitLost: We Tricked GitHub's AI Agent into Leaking Private Repos
Noma Labs discovered a prompt injection vulnerability in GitHub's Agentic Workflows that lets attackers silently extract data from private repos via crafted public issues.
OpenAI's GPT-5.6 Sol Ultra reportedly solves a 50-year-old math problem…
OpenAI bets on families as ChatGPT goes deeper into households
UniClawBench: A Universal Benchmark for Proactive Agents on Real-World…
Terrorist groups are using every major AI chatbot for attack planning…
We charge $10k a week to delete AI-generated code
Slopfix charges $10,000/week to refactor AI-generated codebases that have become unmaintainable — analyzing for free, then cutting code bloat with a committed reduction target.
GPT-5.6 Sol, along with Terra and Luna, will launch publicly this…
OpenAI announces public launch of GPT-5.6 Sol alongside Terra and Luna, possibly satirical naming.
SLORR: Simple and Efficient In-Training Low-Rank Regularization
Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents
Automating AI Away
Re: I'm Begging You to Leave Your AI Note-Taker at Home
AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric…
AI Meets Cryptography 1: What AI Found in Cloudflare's Circl
Latent Memory Palace: Reasoning for Control as Autoregressive…
Workflow as Knowledge: Semantic Persistence for LLM-Mediated Workflows
The Illusion of Equivalency: Statistical Characterization of…
LTM: Large-scale Terrain Model for Wildfire-prone Landscapes
YC CEO says he ships 37K LoC AI code per day. A developer looked under…
Developer investigates YC CEO's claim of shipping 37K lines of AI-generated code daily and looks under the hood.
Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and…
Super Weights in LLMs and the Failure of Selective Training
Beijing is looking at curbing overseas access to China's top AI models
Validity of LLMs as data annotators: AMALIA on authority
Pose-to-Biomechanics: Bridging 3D Human Pose Estimation and…
An agent in 100 lines of Lisp
Deep Learning for Joint Narrowband Interference Cancellation and Soft…
GLM-5.2 (text-generation) by zai-org
GLM-5.2 by zai-org is a multilingual text-generation model with MoE-DSA architecture supporting English, Chinese, and Arabic.
DeepSeek-V4-Pro (text-generation) by deepseek-ai
DeepSeek-V4-Pro by deepseek-ai is a conversational text-generation model with 8-bit/FP8 quantization support.
DeepSeek-R1 (text-generation) by deepseek-ai
DeepSeek-R1 is a reasoning-focused text-generation model using the deepseek_v3 architecture with FP8 support.
Llama-3.1-8B-Instruct (text-generation) by meta-llama
Llama-3.1-8B-Instruct by Meta is an 8B-parameter instruction-tuned model supporting 8 languages.
FLUX.1-dev (text-to-image) by black-forest-labs
FLUX.1-dev by Black Forest Labs is a text-to-image generation model compatible with the diffusers library.
gemma-4-12B-coder-fable5-composer2.5-v1-GGUF (text-generation) by…
Gemma-4-12B-coder-fable5-composer2.5-v1-GGUF is a GGUF-quantized coding model fine-tuned from Google's Gemma-4-12B.
Meta-Llama-3-8B (text-generation) by meta-llama
Meta-Llama-3-8B is Meta's original 8B-parameter base language model from the Llama-3 family.
Llama-2-7b-chat-hf (text-generation) by meta-llama
Llama-2-7b-chat-hf is Meta's 7B-parameter chat-tuned model from the Llama-2 generation.
Meta-Llama-3-8B-Instruct (text-generation) by meta-llama
Meta-Llama-3-8B-Instruct is Meta's instruction-tuned 8B model from the Llama-3 family with Azure deployment support.
bloom (text-generation) by bigscience
BLOOM by BigScience is a 176B-parameter multilingual model supporting 46 languages across 13 families.
gpt-oss-120b (text-generation) by openai
gpt-oss-120b by OpenAI is a 120B-parameter open-source text-generation model with MXFP4 quantization.
gpt-oss-20b (text-generation) by openai
gpt-oss-20b by OpenAI is a 20B-parameter open-source text-generation model with MXFP4 quantization support.
phi-2 (text-generation) by microsoft
Phi-2 by Microsoft is a compact 2.7B-parameter model trained on synthetic 'textbook-quality' data.
stable-diffusion-xl-base-1.0 (text-to-image) by stabilityai
Stable Diffusion XL 1.0 by Stability AI is a high-resolution text-to-image model with 140K+ downloads.
Mistral-7B-Instruct-v0.2 (text-generation) by mistralai
Mistral-7B-Instruct-v0.2 by Mistral AI is a 7B-parameter instruction-tuned model with Apache-2.0 licensing.
Mistral-7B-v0.1 (text-generation) by mistralai
Mistral-7B-v0.1 by Mistral AI is the original 7B-parameter pretrained base model from Mistral.
DeepSeek-V3 (text-generation) by deepseek-ai
DeepSeek-V3 is a conversational text-generation model using the deepseek_v3 architecture with FP8 support.
stable-diffusion-v1-4 (text-to-image) by CompVis
Stable Diffusion v1.4 by CompVis is the original 410M-parameter text-to-image diffusion model.
Llama-3.3-70B-Instruct (text-generation) by meta-llama
Llama-3.3-70B-Instruct by Meta is a 70B-parameter instruction-tuned model supporting 8 languages.
gemma-7b (text-generation) by google
Gemma-7B by Google is a 7B-parameter open-weight model with extensive research paper references.

Research Papers (5 entries)

MPFlow: Learning Budgeted Max-Flow Optimization on the Lightning…
Score Accuracy Along the Forward Diffusion Does Not Certify Numerical…
MulTTiPop: A Multitrack Transcription Dataset for Pop Music
Using AI-based Learning Assistants in Higher Education: A Large-Scale…
Dimensionality Reduction Meets Network Science: Sensemaking on UMAP's…