Skip to content
THE AI WIREINTELLIGENCE THAT MATTERS
Archive

AI news archive

90 articles filed under Chips and compute, newest first, from the outlets listed on the sources page.

All topicsModelsAgentsResearchChips and computeOpen sourceSafetyPolicy and regulationStartups and fundingCommunity
  1. TechRadar AIchips
    RTX 5090s are going for as much as $9000, as AI server builders reportedly buy Nvidia's flagship by the pallet, but the photos behind the story may be AI-generated

    Mainstream AI news from a large tech publication.

  2. WIRED AIchips
    The AI Slowdown Debate Crashed Salesforce’s Party

    The Dreamforce conference became an unlikely battleground for the CEOs of OpenAI, Anthropic, and Nvidia to debate whether AI development should slow down.

  3. AWS Machine Learning Blogchips
    A serverless, data-driven Git metrics dashboard using Amazon Quick Sight

    Learn how to build a fully serverless pipeline that automatically collects Git metrics from GitHub and GitLab and visualizes them in interactive Amazon Quick Sight dashboards, giving engineering teams near-real-time deli…

  4. TechCrunch AIchips
    Huawei plans Q1 2027 launch of new AI chip as it takes on Nvidia

    Huawei is accelerating the launch of its next-generation Ascend 960DT AI chip as it pushes to compete with Nvidia and close China’s AI computing gap with the U.S.

  5. TechCrunch AIchips
    Google, Nvidia, and Anthropic want Emerald AI to find space on the grid for more data centers

    A new coalition that includes Google, Nvidia, Anthropic, and Emerald AI wants to find 100 GW of grid capacity for new data centers.

  6. Engadget AIchips
    NVIDIA and Google's new coalition wants to speed up AI data center power grid connections

    A new coalition of tech giants aims to speed AI data center grid connections in exchange for more flexibility.

  7. TechCrunch AIchips
    Al Gore says the real AI risk isn’t data centers

    In an interview with TechCrunch, Al Gore suggested he isn't losing sleep over AI data center emissions — he's more worried about the AI industry's own warnings about where the technology is headed.

  8. NVIDIA Developer Blogchips
    Serve Qwen3.8-2.4T-A95B, a 2.4T-Parameter Model, with Configurable Reasoning on NVIDIA GB300 NVL72

    Alibaba released the open weights for Qwen3.8-2.4T-A95B (Qwen3.8-Max), its largest open-weight model, bringing near-frontier capabilities to the open...

  9. The Verge AIchips
    The AI data center e-waste problem is huge — and getting bigger

    E-waste from the AI boom has been vastly underestimated, a new report warns. By 2050, it could become enough trash to fill 23 million shipping containers - roughly enough 40-foot containers to circle the world six times…

  10. AWS Machine Learning Blogchips
    Fault tolerant distributed training on Amazon EKS using NVRx

    Integrate NVIDIA Resiliency Extension (NVRx) into PyTorch FSDP training on Amazon EKS to overlap checkpoint I/O with training and recover from GPU faults in seconds. This post covers async checkpointing, in-process resta…

  11. The Verge AIchips
    Apple might make servers again to cash in on the AI rush

    According to The Information, Apple is planning to get back into the server game and might just pair up with Nvidia to make it happen. Apple retired its Xserve line in 2011 and has largely left enterprise machines to oth…

  12. AWS Machine Learning Blogchips
    Build a serverless PII redaction pipeline with Amazon Bedrock Data Automation

    Learn how to automate end-to-end PII detection and redaction from scanned documents at scale using Amazon Bedrock Data Automation with a custom blueprint, AWS Step Functions, and AWS Lambda. A custom blueprint redacts se…

  13. NVIDIA Blogchips
    NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut

    System performance, efficient infrastructure scaling and continuous software optimization are key levers that determine AI inference economics. Higher system performance means more tokens generated, resulting in higher r…

  14. TechCrunch AIchips
    Robots are waiting for a ChatGPT moment: Nvidia’s Les Karpas explains why at TechCrunch Disrupt 2026

    The robotics industry is still waiting for their breakthrough into day-to-day life. Nvidia's Les Karpas has an answer as to why at TechCrunch Disrupt 2026. Register before September 25 to save up to $200 on your pass.

  15. TechCrunch AIchips
    SK Hynix reportedly in talks with Intel to build memory chips in US

    SK Hynix told TechCrunch the company hasn't finalized any plans or arrangements yet.

  16. NVIDIA Blogchips
    Emerald AI, Google and NVIDIA Launch Alliance to Advance Flexible AI Data Centers

    AI factories are the infrastructure of the intelligence era. Scaling them responsibly will depend as much on innovation across the grid as inside the data center. Today, Emerald AI, Google and NVIDIA announced the launch…

  17. NVIDIA Blogchips
    University of Manchester Uses NVIDIA Earth-2 to Forecast Air Pollution Across the UK

    Air pollution is a serious public health risk, contributing to an estimated 30,000 deaths in the U.K. alone last year. Data-driven insights can help — but computing air quality with traditional chemistry-based models is…

  18. NVIDIA Blogchips
    ‘Now We Can Know Everything and Do Anything,’ Jensen Huang Says at Dreamforce

    Know everything. Do anything. That was the message NVIDIA founder and CEO Jensen Huang brought to Salesforce Dreamforce Tuesday, joining CEO Marc Benioff onstage in an appearance that coincided with the announcement of K…

  19. NVIDIA Developer Blogchips
    How NVIDIA Groq 3 LPX Deterministic Execution Drives Power-Efficient High-Interactivity Inference on NVIDIA Vera Rubin

    Power is a defining constraint for AI factories. As AI workloads demand a full compute platform to serve them, each component of that platform must maximize...

  20. NVIDIA Developer Blogchips
    Dense vs. MoE Models: Active Parameters, Throughput, and When to Choose Each

    How can a 30B-parameter model activate only 3B parameters per token, and still use the capacity of the larger model? Nemotron 3.5 Lightning illustrates the...

  21. NVIDIA Blogchips
    From Megawatts to Tokens: How NVIDIA Maximizes AI Factory Production

    On a sweltering August evening in Silicon Valley, as the sun dropped and air conditioning loads spiked, Silicon Valley Power sent a signal to an AI factory to adjust its power consumption. Varun Sivaram was watching on Z…

  22. NVIDIA Blogchips
    AI Infra Summit: NVIDIA Vera Rubin and DSX Platform Advancements Showcase Energy Efficiencies of Optimizing Tokens Per Watt for AI Factories

    Ian Buck, vice president of hyperscale and high-performance computing at NVIDIA, Tuesday spoke on AI factory efficiency at the AI Infra Summit, the Santa Clara Convention Center event that has morphed into a Coachella of…

  23. NVIDIA Developer Blogchips
    How NVIDIA NVLink 6 Delivers Multi-Layer Resiliency for AI Factories

    For operators of large-scale AI factories, maximizing continuous output is essential for productivity. In massive-scale AI training, every GPU in the cluster...

  24. AWS Machine Learning Blogchips
    Optimizing cost and latency with Amazon Bedrock prompt caching

    Prompt caching in Amazon Bedrock can cut input token costs by up to 90% when you repeatedly send the same context to foundation models. This post walks through six practical prompt caching scenarios using the Converse AP…

  25. AWS Machine Learning Blogchips
    Build an AI-powered product tagging system with Amazon SageMaker serverless model customization

    Manually tagging thousands of catalog products is slow and inconsistent. This walkthrough shows how to customize Qwen3-8B with supervised fine-tuning (SFT) and reinforcement learning with verifiable rewards (RLVR) on Ama…

  26. NVIDIA Blogchips
    Heart of the Matter: How a Major Children’s Hospital Uses Open Source NVIDIA AI for Cardiac Care

    AI infrastructure, deployment, and platform coverage from NVIDIA.

  27. NVIDIA Developer Blogchips
    Scaling Federated Learning Across Docker, Kubernetes, and Slurm with NVIDIA FLARE

    Federated learning (FL) projects often begin with a straightforward setup: one server, a few clients, and one dataset at each site. As those projects grow, the...

  28. NVIDIA Developer Blogchips
    Accelerating Dropless MoE Training in JAX with NVIDIA Transformer Engine

    Mixture of experts (MoE) has become one of the defining architectural trends in large-scale AI model training. DeepSeek, Qwen, and Mixtral are examples of MoE...

  29. The Register AI + MLchips
    Teravolt looks to cannibalize older industries to meet AI power demand

    Bitcoin farms, aluminum smelters, and other old infra is more lucrative to repurpose as a datacenter

  30. NVIDIA Developer Blogchips
    From Wafer-Out to First Token: Codifying Supply Chain Expertise with Nemotron and Palantir Foundry

    NVIDIA has one of the largest and most complex supply chains in the world, and its performance is measured from wafer-out to first token. The interval is in two...

  31. NVIDIA Blogchips
    Skild AI Taps NVIDIA Physical AI to Teach Robots New Tasks From a Single Video

    Manufacturing floors, warehouses and production lines rarely stay fixed — tasks change, layouts shift and new products arrive, and most robots can’t keep up without significant reprogramming. Skild AI’s new S1 robot foun…

  32. NVIDIA Blogchips
    d-Matrix Adopts NVIDIA NVLink Fusion for Rack-Scale XPU Deployment

    AI inference chipmaker d-Matrix today announced it will use NVLink Fusion to connect its next-generation Raptor XPUs to NVIDIA’s AI infrastructure platform — joining a growing roster of ecosystem partners. By connecting…

  33. NVIDIA Developer Blogchips
    High-Throughput Structure Prediction with BioNeMo Inference Runtime

    Biomolecular structure prediction is now often run at proteome scale, where the goal is to move an entire worklist through the pipeline efficiently. NVIDIA...

  34. NVIDIA Developer Blogchips
    When to Use Encode-Prefill-Decode Disaggregation to Accelerate Multimodal Model Serving

    Encode-prefill-decode (EPD) disaggregation is an inference optimization technique for multimodal models that separates the vision encoder stage from the prefill...

  35. NVIDIA Developer Blogchips
    CUDA Toolkit 13.4 Adds Windows on Arm Support and Greater Control over Shared GPUs

    Every NVIDIA CUDA Toolkit release adds functionality and performance improvements that help developers get more from NVIDIA GPUs and the broader NVIDIA software...

  36. NVIDIA Blogchips
    NVIDIA Brings Real-Time AI to Broadcast, Sports and Global Streaming at IBC

    At the IBC conference, running Sept. 11-14 in Amsterdam, the creative, technology and business communities are coming together to turn ideas into action and discuss innovations across the media and entertainment industri…

  37. OpenAI Newschips
    GPT-6 Astra: The next generation in intelligence for work

    Meet GPT-6 Astra, OpenAI’s most capable model for business, with advanced reasoning, computer use, and stronger writing and design judgment.

  38. NVIDIA Developer Blogchips
    Introducing CUDA Rust: Two Tracks for Writing GPU Kernels

    In September 2026, NVIDIA announced it is leaning into native GPU programming in Rust. CUDA C++ and CUDA Python are mature, enterprise-grade toolchains, and...

  39. NVIDIA Blogchips
    ‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW

    September is here with 28 more games streaming on GeForce NOW this month, led by a slam dunk: NBA 2K27 with the NVIDIA DLSS 5 3D-Guided Neural Rendering feature. Through NVIDIA’s close collaboration with Visual Concepts…

  40. NVIDIA Blogchips
    NVIDIA to Acquire Hugging Face

    I’m excited to announce that NVIDIA has agreed to acquire Hugging Face for $12,930,300,000. Together, we will scale Hugging Face’s platform, strengthen its infrastructure and expand access to AI for developers and instit…

  41. OpenAI Newschips
    GPT-6 Astra: A new generation of intelligence

    Introducing GPT-6 Astra, our most intelligent and aligned model yet, with state-of-the-art capabilities across computer use, coding, cybersecurity, and science.

  42. NVIDIA Developer Blogchips
    Co-Designing AI Models Using Speculative Decoding for Faster LLM Inference

    This post is the third in a series on AI model co-design. It explores how to accelerate LLM inference while maintaining accuracy using speculative decoding and...

  43. NVIDIA Developer Blogchips
    The Modern CUDA Toolbox in Practice: A Step-by-Step Optimization Walkthrough

    NVIDIA CUDA remains the foundation of GPU-accelerated computing, powering everything from scientific simulations to large-scale AI training. But writing...

  44. Hugging Face Blogchips
    Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI

    Open-source model releases, tooling, and community updates.

  45. NVIDIA Developer Blogchips
    How NVIDIA Groq 3 LPX Unlocks Ultrafast Interactivity at Long Context on NVIDIA Vera Rubin

    NVIDIA Groq 3 LPX is the interactive AI inference accelerator for the NVIDIA Vera Rubin platform. At the core of the platform is NVIDIA Vera Rubin NVL72, the...

  46. NVIDIA Developer Blogchips
    Maximizing AI Factory Performance per Watt with NVIDIA DSX MaxLPS

    AI factories are power-constrained industrial systems. The question is no longer how many GPUs fit in a data center, but how much AI output each available...

  47. NVIDIA Developer Blogchips
    How to Size GPUs for AI Inference and TCO Without Overspending

    The surge in AI adoption is transforming everything from chatbots to content generation. Still, a common pain point remains: How can organizations confidently...

  48. NVIDIA Developer Blogchips
    Scale AV Perception Across Vehicle Platforms with NVIDIA Omniverse NuRec

    A perception stack is shaped by the vehicle that carries it. Move the same software to a new carline—for example, from an SUV to a sedan or another vehicle...

  49. NVIDIA Developer Blogchips
    Deploy an Open Model from Checkpoint to Inference in Two Commands with NVIDIA TensorRT Model Connect

    Open AI models are evolving faster than ever, but bringing them into native applications can still require model-specific conversion, preprocessing,...

  50. OpenAI Newschips
    Supporting Thailand’s next generation of AI startups

    OpenAI and Thailand’s MHESI launch an eight-week accelerator helping 10 health, wellness, and education startups turn AI prototypes into trusted products.