全部 AI 动态
TensorRT Edge-LLM Completes the MLPerf Edge Agentic Benchmark 6.4x Faster on Jetson AGX Thor
AI agents are moving from cloud data centers to vehicles, robots, and other edge devices. Unlike a chatbot that answers a single prompt, an agent works through...
New AI technique could make minimally invasive surgeries safer and more precise
This patient-specific method, called xvr, helps doctors use X-rays for surgical navigation in fields such as orthopedics and neurosurgery.
University of Manchester Uses NVIDIA Earth-2 to Forecast Air Pollution Across the UK
Air pollution is a serious public health risk, contributing to an estimated 30,000 deaths in the U.K. alone last year. Data-driven insights can help — but computing air quality with traditional chemistry-based models is expensive, which limits how detailed they can be and how regularly they can be run. David Topping, a professor in the […]
Measure by measure, studying society accurately
Naoki Egami has become a standout in political methodology, helping refine tools that give scholars durable results.
Gemini Live audio
Tool: Gemini Live audio Google released Gemini 3.8 Live and 3.8 Live Extended Thinking today - two new speech-to-speech models that are a similar shape to OpenAI's GPT-Live family. I pointed GPT-6 Astra Extra High at the documentation and had it build me this web UI for trying out the new models. You can select a model and voice preset, enter an optional system prompt and then start a voice conversation through your browser, including the ability to interrupt the model while it is talking. The implementation uses no libraries. It connects to the wss://generativelanguage.googleapis.com/ws/google.ai.generativelanguage.v1alpha.GenerativeService.
Bypassing inference bottlenecks: Accelerating complex AI search with Retrieve-for-Train
Algorithms & Theory
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
AI for everyone in every language
Animation of several words in different languages slowly zooming past
Heart of the Matter: How a Major Children’s Hospital Uses Open Source NVIDIA AI for Cardiac Care
New method enables AI for safety-critical situations
The “HardFlow” algorithm could help generative AI models produce high-quality outputs that obey strict requirements when “pretty close” doesn’t cut it.
MIT spinout turns plastic waste into resilient building materials
Atlas Building Composites is commercializing MIT research to turn plastic waste into parts for buildings and other infrastructure.
ToolGrad: Efficient tool-use dataset generation with textual "gradients"
Machine Intelligence
How a researcher uses Codex and ChatGPT to search for new antimicrobial molecules
César de la Fuente’s lab uses Codex and ChatGPT to search living and extinct genomes for antimicrobial candidates to fight drug-resistant infections.
Release 5.17.0
Release v5.17.0 New Model additions HYV4 Hy4-Preview is a 780B-parameter mixture-of-experts language model that activates 49B parameters per token. Each MoE layer holds 256 routed experts plus one always-active shared expert and routes every token to 8 of them. The context window is 1M tokens. The architecture combines four features: Multi-head Latent Attention (MLA) compresses keys and values into a low-rank latent ( kv_lora_rank ) that kv_b_proj expands back to one key/value per query head. DeepSeek Sparse Attention (DSA) selects index_topk keys per query with a lightweight indexer. Following IndexShare , only the layers marked "full" in in
MIT Schwarzman College of Computing launches pilot to help educators teach AI across disciplines
A weeklong summer workshop brought higher education faculty to campus to explore how AI and machine learning materials can be adapted for their classrooms.
GPT-6 Astra, Looped Transformers, and Hidden Reasoning
A Look at Recurrent Depth, Hidden Chains of Thought, and Recent Research on Looping Transformer Blocks
Beyond the benchmark: How an adaptive approach drives scientific discovery
For research and development (R&D) organizations, the promise of agentic AI is not a better one-time answer. It is a new way to explore complex scientific and engineering problems: pursuing multiple hypotheses, validating them against evidence, learning from what does not work, and adapting their approach as new information becomes available. The post Beyond the benchmark: How an adaptive approach drives scientific discovery appeared first on Microsoft Azure Blog .
AlphaGenome Atlas: A predictive map of every possible DNA letter change in the human genome
AlphaGenome Atlas maps the molecular effects of 9 billion single-letter DNA variants across the human genome.
Transfer learning for genomic prediction in underrepresented populations
General Science
A connectomics milestone: Mapping the complete male fruit fly brain
General Science
Introducing WeatherNext 3, our most advanced and accurate global weather AI model
NeoMME: an efficient Multimodal-native and Multilingual Encoder
Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps
From MIT to IBM, expediting AI and quantum deployment
MIT affiliates engage with the MIT-IBM Computing Research Lab to bring rigorous theory to production systems.
Introducing Gemini 3.8 Flash and 3.8 Flash Cyber
System helps humans predict when self-driving cars will make mistakes
A new method, called CW-Net, translates the reasoning process of an autonomous vehicle’s AI system into understandable concepts that explain its behavior.
BenchMIRT: What are LLM benchmarks actually measuring?
Walter Torous named executive director of MIT Center for Real Estate
The senior lecturer, already director of the degree program, will now oversee all aspects of the center’s activities and operations.
Mapping global methane emissions from space with deep learning
Climate & Sustainability