Website profile

The New Stack

The New Stack is a media platform for the people who build and manage software the world relies on. We provide context and explanation of at-scale technologies to advance knowledge and create conversations through our coverage of modern architectures, components of the software development life cycle, and operations to

  • 152articles · 30d
  • 8+ hour agolatest article
  • Aug 14, 2026earliest in window
  • 96%with images
  • 88avg words
articles per day
Categories
  • science and technology 140
  • SOD 101
  • CE 95
  • NW 38
  • SOF 21
  • SCT 13
  • economy business and finance 9
  • FIB 9

Please confirm you are human

This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.

A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.

Hold with a pointer, or hold Space or Enter.

News

The New Stack
thenewstack.io > openai-jalapeno-inference-chip

OpenAI built a chip in nine months. Then it let AI rewrite the code.

2+ week, 5+ day ago   (423+ words) OpenAI published its first Jalapeño chip benchmarks, showing big gains in throughput and latency that target the compounding delays AI agents create....

The New Stack
thenewstack.io > ai-factories-are-among-the-most-complex-systems-ever-built-nvidia-and-palantir-turn-nvidias-supply-chain-into-a-proving-ground-for-sovereign-ai

“AI factories are among the most complex systems ever built”: Nvidia and Palantir turn Nvidia’s supply chain into a proving ground for sovereign AI

3+ day, 14+ hour ago   (188+ words) makes supply chains a natural target for the technology. From chips and memory to manufacturing, networking,...

The New Stack
thenewstack.io > nvidia-open-models-chips

Nvidia is paying $12.9 billion to keep open models on its chips

2+ week, 2+ day ago   (334+ words) Ollama put Qwen, DeepSeek and Kimi inside Claude Desktop. Nvidia’s reported Hugging Face acquisition is a bet that the open-model ecosystem will keep running on CUDA....

The New Stack
thenewstack.io > cut-gpu-cold-starts

Cut GPU inference cold start from 8 minutes to less than a minute

1+ week, 3+ day ago   (635+ words) Cut GPU inference cold start times from 8 minutes to under 30 seconds with simple configuration and platform fixes....

Web

External web results are waiting for the human check. Complete the press-and-hold control above. Google advertising and AI choices remain separate after verification.