Skip to content
THE AI WIREINTELLIGENCE THAT MATTERS
Archive

AI news archive

997 articles, newest first, from the outlets listed on the sources page.

All topicsModelsAgentsResearchChips and computeOpen sourceSafetyPolicy and regulationStartups and fundingCommunity
  1. OpenAI Newsmodels
    Introducing the Admin plugin for ChatGPT Work and Codex

    Use the Admin plugin for ChatGPT Work and Codex to analyze workspace usage, manage members and permissions, adjust limits, and act on admin requests.

  2. Hugging Face Blogagents
    Wire It, Run It, Deploy It: AI Workflows in Gradio

    Open-source model releases, tooling, and community updates.

  3. NVIDIA Developer Blogchips
    CUDA Python 1.0: Stable APIs, One Foundation, Full Platform Access

    For years, a Python developer who needed a GPU had two realistic choices: Learn NVIDIA CUDA C++ well enough to write an extension, set up a build toolchain, and...

  4. Mistral AI Newsmodels
    Mistral x HUMAIN

    Open frontier models, product updates, and developer tooling from Mistral AI.

  5. NVIDIA Developer Blogagents
    NVIDIA Vera Rubin and Blackwell Set a New Standard for Agentic AI Performance per Watt

    AI agents have expanded inference from single-turn interactions into multi-step workflows that reason, invoke tools, coordinate subagents, and carry growing...

  6. NVIDIA Developer Blogagents
    NVIDIA BlueField-4 Powers New Scale-In Network Infrastructure for Agentic AI Factories

    Traditional cloud infrastructure was designed for predictable, general-purpose workloads and standard interfaces. Agentic AI factories connect diverse users,...

  7. OpenAI Newsmodels
    Advancing price-performance for developers with GPT‑5.6 in Kiro

    GPT‑5.6 is now available in Kiro, helping developers plan, build, review, and test software with better price-performance.

  8. Guardian Technologycommunity
    Fairphone 6+ review: the most repairable, ethical phone gets faster

    Dutch sustainable smartphone revamped with new chip and more RAM without sacrificing modular design Fairphone’s top Android for 2026 is a little faster with more memory while maintaining full compatibility with its previ…

  9. NVIDIA Developer Blogagents
    Solving Agentic AI Fleet Challenges with NVIDIA Vera CPU

    AI factories are interconnected systems where fleet economics depend on how efficiently the entire stack converts power and capital into completed agent tasks....

  10. NVIDIA Developer Blogagents
    Where Security Fits in an AI Agent Stack

    As AI agents become more capable and operate over longer horizons, building security and trust into the applications they power becomes increasingly important....

  11. NVIDIA Developer Blogagents
    NVIDIA AVO Reaches 100% on ARC-AGI-3, Demonstrating a Frontier-Level General-Purpose Architecture for Long-Horizon Autonomous Agents

    A frontier language model is only one component of an AI agent. The surrounding agent system—often called a harness—determines how the model receives...

  12. NVIDIA Developer Blogmodels
    How Generative Recommenders Are Redefining RecSys at Scale

    Recommender systems (RecSys) are one of the most ubiquitous machine learning problems in the consumer internet industry yet notoriously difficult to train and...

  13. NVIDIA Developer Blogchips
    GPU-Accelerated Clustering for Financial Instruments at Scale

    Use AdaptGrow, a GPU-accelerated matrix factorization algorithm, to turn rolling correlation and tail-dependence matrices into hard clusters, soft factor...

  14. Google DeepMindresearch
    From Atari to EVE Online: Building on 15 Years of AI Research in Games

    Google DeepMind partners with game studios to prototype breakthrough AI gameplay.

  15. Ai2 Blogmodels
    How a Georgia Tech team used the open Olmo stack to trace social reasoning

    A Georgia Tech team used Ai2’s fully open Olmo stack to trace social reasoning back to the training data that shaped it, finding that dialogue-rich, interpersonal writing had an outsized influence on the capability.

  16. Hugging Face Blogresearch
    How Hugging Face Inference Endpoints, Jobs, and Buckets Power Search on Papers with Code

    Open-source model releases, tooling, and community updates.

  17. Hugging Face Blogresearch
    Measuring benchmark optimization in speech recognition

    Open-source model releases, tooling, and community updates.

  18. NVIDIA Developer Blogagents
    NVIDIA Nemotron 3.5 Lightning Delivers Fast, Accurate Specialized Task Execution for Long-Running Agents

    Long-running AI agents spend most of their time on high-volume execution: tool calls, result validation, and subagent delegation. Using a frontier reasoning...

  19. NVIDIA Developer Blogagents
    NVIDIA Nemotron 3 Ultra Leads Open Models on Accuracy and Efficiency in Agentic RTL Coding

    Modern chip design is increasingly limited by engineering time. Register transfer level (RTL) development and verification require specialized hardware...

  20. NVIDIA Developer Blogopen-source
    ModelExpress: Distributing Model Artifacts at the Speed of Light

    Every byte moved has a cost. As model checkpoints grow to hundreds of gigabytes or even a terabyte, that cost adds up quickly. To make things even worse, moving...

  21. NVIDIA Developer Blogagents
    Six Agent Harness Capabilities for Higher Model Performance

    Building a great AI agent isn’t just about choosing the right models. The harness is the architecture surrounding the model. How it renders context, executes...

  22. NVIDIA Developer Blogchips
    Advancing Semiconductor Innovation Across Materials Engineering and Manufacturing

    As AI workloads increase, explosive compute demand is pushing the semiconductor industry to meet unprecedented performance targets. Even small delays can have...

  23. NVIDIA Developer Blogagents
    Developing Healthcare Robotics with GPU-Native Medical Physics Simulation

    Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation....

  24. NVIDIA Developer Blogchips
    NVIDIA Ising Enables Fully Automated Quantum Computer Calibration with Enhanced In-Context Learning

    NVIDIA Ising Calibration is an open source vision language model (VLM) designed to interpret diagnostic outputs from quantum processors and determine how they...

  25. NVIDIA Developer Blogchips
    NVIDIA Exemplar Cloud: Lessons for Unlocking Full Performance on AI Infrastructure

    Two AI computing clusters built from identical NVIDIA H100, GB200 NVL72, or GB300 NVL72 systems can deliver materially different training throughput. We...

  26. NVIDIA Developer Blogagents
    How to Self-Host a Validated AI Coding Assistant with NVIDIA NeMo Guardrails

    Deploying an AI coding assistant in a regulated, sovereign, or source-sensitive environment, often comes with challenges. Three common issues are: the source...

  27. NVIDIA Developer Blogchips
    Run High-Performance Core Math at Scale with NVIDIA nvmath-python

    NVIDIA nvmath-python is a library designed to bridge the gap between the Python scientific community and NVIDIA CUDA-X math libraries. It gives Python users...

  28. NVIDIA Developer Blogagents
    Four Ways to Deploy More Secure AI Agents

    Knowledge workers are increasingly integrating AI agents into their workflows. Agents that function as "digital coworkers" offer clear benefits. For example,...

  29. NVIDIA Developer Blogagents
    Co-Designing AI Model Attention for Fast, Interactive Long-Context Inference

    As agentic and long-context workloads become common, the context lengths increase and attention consumes a larger share of inference time (Figure 1). Because...

  30. NVIDIA Developer Blogchips
    NVIDIA Video Codec SDK 13.1: Zero-Copy Transcode, AV1 B-Frames, and Frame-Accurate Seek

    The demand for high-quality video continues to accelerate across industries, powering everything from immersive streaming experiences to remote collaboration,...

  31. NVIDIA Developer Blogchips
    How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure

    Running a dedicated Kubernetes cluster per team often results in more isolation than an organization requires. While one cluster can be successfully shared...

  32. NVIDIA Developer Blogagents
    NVIDIA Vera Storage Benchmarks: Faster Encryption, Compression, Integrity Checking, and Recovery for AI-Native Storage

    Storage is an active part of every agentic AI workflow. As agents retrieve enterprise knowledge, access persistent memory, reuse key-value (KV) cache data,...

  33. NVIDIA Developer Blogagents
    Generate Trajectories, Reasoning Traces, and Auto-Labels with NVIDIA Alpamayo 2 Super

    Autonomous vehicle (AV) development often relies on separate models for trajectory generation, high-level intent prediction, scene understanding, and data...

  34. NVIDIA Developer Blogpolicy
    Beyond VLAs: How World Action Models Reshape Robot Manipulation

    A central challenge in robotics is building policies that generalize beyond the demonstrations they’re trained on. A policy that succeeds in a training scene...

  35. NVIDIA Developer Blogagents
    Route AI Agent Workloads Across Models with NVIDIA NeMo Switchyard

    Building an AI agent does not end with choosing a single model. Each model has its own strengths, weaknesses, and cost profile, which can shift from one...

  36. NVIDIA Developer Blogchips
    How to Choose Full-Stack Observability for NVIDIA AI Factories

    AI infrastructure spans multiple layers, from compute and networking to storage, orchestration, and applications. When performance degrades, identifying the...

  37. NVIDIA Developer Blogagents
    NVIDIA JetPack 7.2.1 Adds Agentic Video Skills and T3000 Emulation

    Video is a core data path across NVIDIA Jetson applications, from robotics and intelligent video analytics to industrial automation, healthcare, media...

  38. NVIDIA Developer Blogchips
    Run Massive-Scale UMAP in Minutes Using Multiple GPUs—Without Losing Accuracy

    Uniform Manifold Approximation and Projection (UMAP) is a dimensionality reduction technique widely used for visualization and feature extraction. Applications...

  39. NVIDIA Developer Blogchips
    Developing Nemotron 3.5 Lightning NVFP4 with QAD Using NVIDIA Model Optimizer

    Teams customize their models to hit their targets for latency, speed, memory, and compute. With the open NVIDIA Nemotron family of models, developers can find...

  40. NVIDIA Developer Blogchips
    Post-Train NVIDIA Cosmos 3 Edge for On-Device Robot Control

    Robots need policies that can adapt to their sensors, environments, and tasks while running on onboard computing hardware. World models offer a foundation for...

  41. NVIDIA Developer Blogagents
    Building Federated Multimodal AI Workflows with NVIDIA FLARE

    Modern vision-language models (VLMs) can support tasks such as visual question answering, captioning, and image-text reasoning. In practice, however, the data...

  42. NVIDIA Developer Blogagents
    Evaluating AI Agent Skill Performance with NVIDIA SkillEvaluator

    AI agents are only as effective as the context they receive. Even with capable models and well-documented NVIDIA libraries, agents can spend extra steps finding...

  43. NVIDIA Developer Blogagents
    Developing NVIDIA Holoscan Applications with CLI, Skills, and AI Coding Agents

    NVIDIA Holoscan is a platform for building real-time AI applications at the edge, from medical imaging to robotics. HoloHub is its companion repository: a...

  44. Hugging Face Blogchips
    Up to 3.2x Faster Inference with LFM2.5-DSpark

    Open-source model releases, tooling, and community updates.

  45. Microsoft Researchresearch
    Broadening access to Skala creates a faster path to predictive DFT

    Skala 1.1, the updated deep-learning exchange-correlation functional from Microsoft Research, provides greater accuracy, expanded accessibility across the computational chemistry ecosystem, and a living benchmark to trac…

  46. Mistral AI Newsagents
    Agentic Search. More accurate and efficient results from your AI systems.

    The retrieval layer that helps AI systems navigate, read, and verify information inside even the most complex documents

  47. OpenAI Newspolicy
    Introducing Intelligence Age

    Introducing Intelligence Age, a new OpenAI blog exploring how transformative AI could reshape power, governance, the economy, and individual freedom.

  48. OpenAI Newsmodels
    Stampli cuts launch hours by 68% using ChatGPT Work

    With a fixed deadline and design resources committed elsewhere, Stampli used Codex and ChatGPT Work to compress weeks of launch production into days.

  49. OpenAI Newssafety
    Offering Zero Data Retention for frontier models

    OpenAI reaffirms Zero Data Retention for eligible API customers and previews Private Safety Processing for advanced AI safety without compromising data privacy.

  50. Google AI Blogresearch
    5 new ways to level up your learning with Search

    Here’s how you can use Google Search tools to study for classes and standardized tests.