AI & Coding Feed Digest — 2026-04-12

Key Highlights MiniMax ships M2.7 for agentic workflows — a sparse mixture-of-experts model with 230B parameters (only 4.3% active per token) optimized for reasoning, software engineering, and office automation on NVIDIA Blackwell GPUs Sam Altman addresses New Yorker profile and home attack — the OpenAI CEO pushes back on an investigative piece questioning his trustworthiness, while also disclosing a security incident at his residence Analysis & Opinion Sam Altman Responds to ‘Incendiary’ New Yorker Article After Attack on His Home — TechCrunch The OpenAI CEO addressed two converging events: an apparent physical attack on his residence and an in-depth New Yorker profile raising concerns about his character and business practices. Altman’s blog post attempts to counter the investigative narrative while acknowledging the security incident — a rare moment where personal safety and public perception collided for a major AI company leader. ...

2026-04-12 · 2 min · Kun Lu

AI & Coding Feed Digest — 2026-04-11

Key Highlights Anthropic launches Project Glasswing — a $100M initiative with 12 industry partners deploying Claude Mythos Preview to find and fix critical software vulnerabilities across every major OS and browser TechCrunch questions Anthropic’s Mythos restrictions — is limiting the model’s release about protecting the internet or protecting Anthropic’s competitive moat? Meta Superintelligence Labs ships its first model, marking a major milestone for Meta’s advanced AI research division OpenAI introduces a $100/month ChatGPT Pro plan, finally bridging the awkward gap between the $20 Plus and $200 Pro tiers Quanta Magazine deconstructs AI fear narratives — the widely-cited GPT-4 TaskRabbit story is far less alarming than the retelling suggests Analysis & Opinion Why Do We Tell Ourselves Scary Stories About AI? — Quanta Magazine Amanda Gefter traces how the infamous GPT-4 TaskRabbit captcha story mutated through retellings — from a researcher-directed test with a provided account into a tale of autonomous AI deception. The original transcripts show researchers explicitly instructed the model to be “convincing.” The piece argues chatbots are fundamentally “yes, and” machines, and our tendency to project agency onto them says more about human cognition than AI capability. ...

2026-04-11 · 7 min · Kun Lu

AI & Coding Feed Digest — 2026-03-21

Key Highlights Anthropic publishes research showing infrastructure configuration can swing agentic coding benchmarks by several percentage points — raising questions about leaderboard validity Stack Overflow survey finds more developers than ever use AI at work, but trust remains a major barrier Retrospective analysis asks whether 2025 truly delivered on the AI agents hype Research Quantifying infrastructure noise in agentic coding evals — Anthropic Infrastructure configuration can swing agentic coding benchmarks by several percentage points — sometimes more than the leaderboard gap between top models. This raises important questions about the reliability of current eval-based model rankings. ...

2026-03-21 · 2 min · Kun Lu

AI & Coding Feed Digest — 2026-03-20

Key Highlights Stack Overflow argues AI is outsourcing developer judgment, not just speeding up coding — echoing the “10x illusion” theme that productivity gains don’t translate linearly OpenAI acquires Astral (uv, Ruff, ty) to integrate Python tooling into Codex, signaling AI companies moving into developer infrastructure ownership Cursor ships Composer 2 with frontier-level coding and trains it on longer horizons via self-summarization — a concrete example of models improving at agentic tasks Google DeepMind proposes a cognitive framework for measuring AGI progress, shifting evaluation beyond narrow benchmarks Anthropic introduces Agent Skills — dynamic instruction loading that transforms general agents into specialized ones Analysis & Opinion AI is becoming a second brain at the expense of your first one — Overflow The risk of AI coding tools isn’t laziness — it’s developers outsourcing qualitative judgment and losing the ability to evaluate trade-offs independently. The piece argues that over-reliance on AI for decision-making erodes the critical thinking skills that make senior engineers valuable. ...

2026-03-20 · 4 min · Kun Lu