Skip to main content

AI News

91 items

  1. OpenAI

    Better prompt caching for GPT-6

    This note describes how GPT-6 improves prompt caching: higher cache hit rates, new diagnostics, explicit breakpoints, and controls aimed at reducing latency and costs.
  2. OpenAI

    Introducing GPT-6 Sol and Luna: Two Frontier Models for Everyday Work

    The piece announces two models, GPT-6 Sol and Luna, described as bringing frontier intelligence to everyday work while being differentiated by different balances of capability and cost.
  3. IEEE Spectrum

    The Future Is Fanless: 100% Heat Capture for Liquid Cooled AI Servers

    This CoolIT-provided article argues that beyond roughly 250 kW per rack a 70/30 liquid-to-air split leaves a 75 kW air load, so near-total liquid heat capture with air below 1% enables fanless AI servers; it describes heat cascading from processors into memory, networking, storage, and power components, CoolIT's modular coldplates plus conductive plates, vapor chambers, heat pipes, thermal transfer plates, and riding coldplates unified in one server loop, and its modeling that places full heat capture as standard for flagship rack-scale products through 2028.
  4. MIT Technology Review

    The Download: Why AI's Latest Breakthroughs and Fears May Be More Hype Than Reality

    This edition of The Download centers on a commentary by Timnit Gebru and Emily M. Bender arguing that recent breathless claims about AI hacking, mathematical breakthroughs, and the prospect of self-improving superintelligence tend to look very different once experts examine what happened, and that there is a strong commercial incentive to overstate AI capabilities, with the illusion of speed and urgency also serving to misdirect policymakers and the public; the newsletter also rounds up the day's technology stories, including 22 nations calling for a new global AI oversight body, Texas and California moving to rein in data centers, AMD becoming the twelfth $1 trillion company, Meta's AI agent Muse topping the US App Store, an edible battery, political risk to NIH grants, epigenetic editing
  5. NVIDIA Research

    NVIDIA Isaac ROS 5.0 Advances Agentic, Open Source Robotics Development

    NVIDIA released Isaac ROS 5.0 at ROSCon Toronto, a collection of GPU-accelerated packages built on ROS that adds support for ROS Lyrical and Ubuntu 24.04, contributes a standard data-handling interface with the Open Source Robotics Alliance, introduces agent-ready Isaac skills and documentation, a FoundationStereo fine-tuning skill, an agent-ready FoundationPose inference library for object pose estimation and tracking with up to 5.5x faster performance, and a standalone pick-and-place skill, while showcasing ecosystem integrations such as RealSense, Intrinsic, Seeed Studio, Magna, and Ekumen across NVIDIA Jetson Orin Nano to Jetson Thor platforms for perception, navigation, and manipulation.
  6. OpenAI

    Parallel Halved Research Time and Cost with GPT‑6 Astra

    Parallel's agents used GPT‑6 Astra to research and synthesize labor-market data, and according to this text the result was halving both research time and cost relative to prior models.
  7. Hugging Face

    Transformers now runs llama.cpp quants

    This Hugging Face post describes how transformers can now load GGUF quantized checkpoints directly through from_pretrained with a gguf_file argument: on Apple Silicon it reuses ggml Metal kernels distributed via the kernels library (packed-quantized weight matmul, fused normalization, flash attention, gated delta net, plus a home-grown topk for MoE routing) and trims generate overhead by dropping the redundant attention mask early and deferring the stopping check asynchronously; on a MacBook Pro M2 Max the authors measure generation speed close to llama.cpp across three checkpoints, report file-size tradeoffs for Q4_K_M, Q5_K_M and Q6_K, and state that packed inference is MPS-only with architecture coverage for Qwen3.5 dense and MoE (including compatible Qwen3.
  8. Amazon Science

    Amazon and Stanford Launch Joint Research Initiative to Advance AI and Science

    Amazon announced the Stanford and Amazon Research Initiative with Stanford University, a formal framework to advance research at the frontiers of AI, energy, and healthcare through joint research projects, PhD fellowships, and symposia, aiming to move breakthrough research into real-world solutions faster and broaden participation from diverse scholars.
  9. NVIDIA Research

    NVIDIA Launches DSX Ready: A Qualification Program for Power and Cooling Products in AI Factories

    NVIDIA introduced DSX Ready, a qualification program for partner products and solutions that meet applicable NVIDIA DSX AI factory reference design requirements, launching with two initial categories, battery energy storage systems (BESS) and cooling distribution units (CDUs), and naming the first qualified BESS and CDU suppliers.
  10. NVIDIA Research

    Why Deploying Physical AI at Scale Demands Safety at Every Layer

    This article explains why physical AI—autonomous vehicles, humanoid robots, and industrial robots—needs safety spanning hardware, software, AI behavior, operating environment, and the deployment lifecycle as it moves from research to large-scale deployment; it presents NVIDIA Halos as a full-stack safety system across AV and robotics lines (including DRIVE AGX Thor, Hyperion, Halos OS, Alpamayo, IGX Thor, Holoscan Sensor Bridge, Isaac Lab, and Omniverse), and lists the automakers, robotics firms, chip and sensor suppliers, and certification bodies in that ecosystem, along with third-party assessments and accreditations by TÜV SÜD, TÜV Rheinland, and ANAB.

Page 7 · showing 10