Etched's Sohu Chip: A Transformer-Specific Bet That Just Raised $800 Million

Etched's Sohu Chip: A Transformer-Specific Bet That Just Raised $800 Million

Etched emerged from stealth today with $800 million in funding and over $1 billion in signed customer contracts, unveiling its Sohu chip – an ASIC designed exclusively for transformer inference. The move signals a growing push to build domain-specific hardware for large language models, potentially reshaping the AI chip market.

July 1, 2026 · 3 min · 516 words · Rajesh
OpenAI's Codex Has Quietly Taken Over – 99.8% of Internal Tokens and a 137x Surge in Non-Developer Use

OpenAI's Codex Has Quietly Taken Over – 99.8% of Internal Tokens and a 137x Surge in Non-Developer Use

OpenAI’s Codex dominates internal usage with 99.8% of weekly output tokens, while non-developer adoption surges 137x. A look at what this means for AI tooling.

June 30, 2026 · 3 min · 518 words · Rajesh
Reflection AI Bets $6.3 Billion on SpaceX for Open-Source Frontier AI

Reflection AI Bets $6.3 Billion on SpaceX for Open-Source Frontier AI

Reflection AI signed a $6.3B deal with SpaceX for GB300 chips at Colossus 2 through 2029, aiming to end reliance on proprietary AI like Anthropic and OpenAI.

June 29, 2026 · 3 min · 439 words · Rajesh
OpenAI GPT-5.6 Launches a Three-Model Family: Sol, Terra, and Luna

OpenAI GPT-5.6 Launches a Three-Model Family: Sol, Terra, and Luna

OpenAI has launched GPT-5.6, a family of three models, with a cautious rollout to trusted partners first. The flagship Sol features Max and Ultra modes for deep reasoning.

June 28, 2026 · 2 min · 316 words · Rajesh
NVIDIA AI Fixed Its Own Broken Metric While Training a 30B Model—Autonomously

NVIDIA AI Fixed Its Own Broken Metric While Training a 30B Model—Autonomously

An autonomous AI system trained a 30B parameter model and self-corrected its evaluation metric—a first at frontier scale.

June 27, 2026 · 3 min · 427 words · Rajesh
IBM's Sub-1nm Chip Breaks Silicon Limits, Promises 50% Performance Leap

IBM's Sub-1nm Chip Breaks Silicon Limits, Promises 50% Performance Leap

IBM announces a 0.7nm chip with 100 billion transistors, doubling density over its 2nm node and delivering up to 50% more performance or 70% lower power.

June 26, 2026 · 3 min · 496 words · Rajesh
Abstract illustration of an AI agent orchestrating tools and subagents

Building a Custom Agent with the Claude Agent SDK

A practical, code-first walkthrough of the Claude Agent SDK: custom tools, MCP servers, and subagents — the same engine that powers Claude Code, in your own app.

June 26, 2026 · 6 min · 1177 words · Rajesh
Fable 5 Writes a Bootable Windows Kernel in Rust – AI Now Writes Critical Infrastructure Code

Fable 5 Writes a Bootable Windows Kernel in Rust – AI Now Writes Critical Infrastructure Code

Anthropic’s Claude Fable 5 generated a bootable Windows NT kernel in Rust in 38 minutes, passing all 14 self-tests. The feat demonstrates AI’s ability to write core OS software, sparking security debates.

June 25, 2026 · 3 min · 441 words · Rajesh
Human-Like Memory Limits Make AI Better at Learning Grammar

Human-Like Memory Limits Make AI Better at Learning Grammar

A new study demonstrates that small language models with transient memory learn grammar more efficiently than massive models, suggesting limitations, not scale, are key to language acquisition.

June 24, 2026 · 3 min · 453 words · Rajesh
OpenAI Codex Can Now Watch and Learn Your Workflow — No Coding Required

OpenAI Codex Can Now Watch and Learn Your Workflow — No Coding Required

OpenAI’s Record and Replay lets Mac users demonstrate a workflow once and generate a reusable Codex skill. No coding skills needed.

June 23, 2026 · 3 min · 453 words · Rajesh