Hello world, it’s Thursday, September 10th.
DOJ is looking into Nvidia-Groq deal, d-Matrix Raptor chips to work alongside NVIDIA racks, DeepSeek gets way way cheaper, OpenAI makes research intern → goes for researcher’s job next, and a real mysterious memory company.
Let’s get into it. — Austin & Vik
Be sure to check out the Semi Doped podcast on YouTube or your favorite podcast player!
Tell us how to make Semi Doped better! Give us your feedback! ❤️
DOJ Probes Whether Nvidia Structured Groq Deal to Skirt Merger Review
The Justice Department has opened an antitrust investigation into Nvidia’s licensing deal with AI chip startup Groq, according to reporting cited by Reuters. The probe centers on whether Nvidia deliberately framed the arrangement as a licence rather than an acquisition to stay below thresholds that would trigger mandatory merger review. The deal is valued at roughly $17 billion to $20 billion, depending on the source, making it substantial enough that its structure, not just its size, drew federal attention. Nvidia’s stock dipped on the news. The DOJ has not filed any action, and neither company has publicly commented on the investigation’s scope.
Vik: The acqui-hire deal was exactly supposed to avoid an anti-trust, but DoJ hot on Nvidia’s case after all. Guess the plan didn’t work out as expected? Let’s see how it goes.
Austin: To be fair, there was precedent before Nvidia. Remember Microsoft + Inflection, Google + Character.AI?
d-Matrix Ties Raptor XPU to NVLink Fusion and Nvidia MGX Rack Standard
Inference chipmaker d-Matrix announced it will connect its next-generation Raptor XPUs to Nvidia’s AI infrastructure through NVLink Fusion, integrating with both NVLink scale-up networking and the Spectrum-X scale-out fabric. The company also adopts the Nvidia MGX rack architecture, meaning its hardware ships inside a chassis designed to Nvidia’s own mechanical and electrical spec. d-Matrix is joining what Nvidia describes as a growing roster of ecosystem partners, a list that now spans custom silicon startups willing to wire their products into Nvidia’s interconnect stack rather than build independent rack topologies.
Austin: If you’re an AI accelerator startup you COULD build everything — compute, scale-up, cooling, rack design, etc… or you could partner with Nvidia and use their scale-up networking, cooling etc which is already deployed in production worldwide…
Vik: Raptor gains access to Nvidia AI factory deployments, and Nvidia gains another accelerator vendor whose product only fully performs inside its own infrastructure. This is good news for d-Matrix whose 3D DRAM approach I personally like a lot for AI inference. It’s a challenging engineering problem, but even Qualcomm’s HBC is along the same vein.
DeepSeek’s V4.1 Flash Model Prices Inference at Fractions of a Cent
DeepSeek released V4.1 Flash on Thursday, built on a 552-billion-parameter Mixture-of-Experts architecture using what the company calls a “Causal-Encoder-Decoder” design. The model charges as little as a fraction of a cent per million tokens, undercutting rivals including Anthropic, OpenAI, and Z.AI. DeepSeek also claims V4.1 Flash beats Kimi K3 on cybersecurity and coding benchmarks while running faster than its previous flagship. The pricing isn’t accidental; each successive DeepSeek release has arrived cheaper and more capable, compressing the global floor for inference. Hyperscalers that built CapEx cases around premium model pricing now face a competitor content to charge almost nothing.
Vik: DeepSeek has been continuously improving on how they compress KV cache, which means they need to use even lesser SSD offload. The progress is actually impressive. zephyr_z9 on X posted a good picture of this.
OpenAI Reports Automated Research Intern Goal Met, Targets Full AI Researcher by 2028
OpenAI published internal metrics on September 6, 2026 showing its research organization now uses 3.1 agent-workdays of effort for every workday of human labor, up from below parity before June 2026. The company states it has met its previously announced goal of building an automated research intern by September 2026, defined as a system that can complete well-defined tasks under human direction that would take a skilled researcher a few days. The company says it is targeting an automated AI researcher by March 2028 and plans to continue public reporting on its progress toward recursive self-improvement.
Vik: ✅ AI intern check, 📌 AI researcher. 3.1 agent workdays / 1 human workday - so thats the metric here, and its rapidly accelerating. By the end of the year, we should easily have a 10x employee that never tires. Is any job safe now?
Austin: Vik, when are you creating a Semi Doped intern for us?
Wafer Launches AI Performance Engineering Series Starting with Prefill vs Decode
Wafer announced the launch of what it describes as the most comprehensive AI performance engineering repository, followed by a deep-dive series covering each resource. The series begins with Part 1 on prefill vs decode.
Sources: @wafer_ai
Vik: Great learning resource, highly recommend.
👀 Look at this mysterious memory company
See more at keplercompute.com.
Sector Watch
Foundry & Packaging
GlobalFoundries and Monolithic Power Systems sign a long-term manufacturing agreement to produce advanced power solutions at GF’s Singapore fab. (quiverquant.com)
Samsung Foundry and Cadence expand their AI design ecosystem partnership to cover next-generation AI infrastructure and physical AI process nodes. (Samsung Semiconductor)
Taiyo Holdings launches the FPIM Series packaging material, achieving a three-layer RDL formation with 700 nm critical dimensions on 12-inch wafers for advanced semiconductor packaging. (PR Newswire)
Memory
UMC reports August revenue up 30%, as mature-node demand tightens. (inkl)
Huawei raises the price of its Ascend 950DT AI accelerator by 60% over three months; Cambricon also lifts prices as domestic HBM supply remains critically short. (Bloomberg.com)
Simmtech unveils a next-generation AI semiconductor substrate technology aimed at advanced packaging demand from high-bandwidth memory stacks. (아시아경제)
Compute & Edge
AMD ships 50,000 MI450 GPUs from its Helios generation to Oracle, marking the first large-scale deployment of the new accelerator. (shattered.io)
Qualcomm details its next Snapdragon flagship with an upgraded Hexagon NPU capable of running 30-billion-parameter on-device AI models. (ServeTheHome)
Renesas launches the 64-bit RZ/G3L and RZ/G3SE MPUs targeting HMI and IoT edge designs with a unified development environment. (EE News Europe)
China Chip Industry
Enflame debuts on the Hong Kong Stock Exchange following its $913 million IPO, completing China’s ‘four-dragons’ domestic GPU public listings. (digitimes)
Loongson raises $340 million to expand homegrown CPU and GPU development as AI inference demand boosts domestic chip plans. (digitimes)
SMIC posts quarterly revenue above $3 billion for the first time, narrowing the gap with Samsung Foundry to roughly $250 million. (XenoSpectrum)
Power
DB HiTek completes reliability qualification for its 8-inch, 1,200V SiC MOSFET process and targets mass production in 2027, claiming the world’s first 8-inch SiC foundry line. (Chosunbiz)
Infineon and SolarEdge extend their partnership to deliver solid-state DC fault protection for 800 VDC AI data-center power architectures. (Energetica India Magazine)
ASML and Xanadu announce a collaboration to optimize lithography processes for ultra-low-loss quantum photonic hardware. (Quantum Zeitgeist)
Data Centers & Infrastructure
Wistron and Wiwynn both hit record August revenues; Wiwynn says AI server shipments will overtake general-purpose servers in Q3 2026 and is adding a 120 MW power plant for its Socorro factories. (digitimes)
Nvidia announces a major expansion of data-center capacity in Australia, partnering with local operators to meet surging AI compute demand. (Nvidia News)
Crux AI, the Google-Blackstone TPU neocloud venture, formally launches and hires Meta’s former data-center engineering head Alan Duong. (Data Center Dynamics)
Optics & Networking
Huawei unveils a new high-speed optical module and proprietary optical interconnect standard positioned as a direct challenge to Nvidia and Broadcom in AI infrastructure. (Nikkei Asia)
Singtel and Gulf Development forge a strategic partnership to build a new subsea cable link between Singapore and Thailand. (Light Reading)
Spark Wholesale launches a dedicated data-centre interconnect service in Auckland built on Ciena optical transport equipment. (reseller.co.nz)
Tell us how to make Semi Doped better! Give us your feedback! ❤️





