Loading…

AI & Machine Learning News Hub

Research, releases, and applied work in AI & ML

Latest
Quanta MagazineNeutrinos From Deep Inside Earth Provide a New Picture of the MantleAi2 BlogTutorMoments: Do AI tutors know when to help and when to hold back?cs.LG updates on arXiv.orgMS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classificationcs.LG updates on arXiv.orgWhen Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasterscs.LG updates on arXiv.orgPPDL: LLM-Based Flows as Probabilistic Programscs.LG updates on arXiv.orgDecoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Languagecs.LG updates on arXiv.orgDisentangling 3D Modeling from Spatial Reasoningcs.LG updates on arXiv.orgMarginal Matching Does Not License Factorized Sampling: Auditing Conditional Style Leakage in Factorized Generative Modelscs.LG updates on arXiv.orgPRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesiscs.LG updates on arXiv.orgBeyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuningcs.LG updates on arXiv.orgBeyond Rotations: AuroOFT for Expressive Quantized Orthogonal Fine-Tuningcs.LG updates on arXiv.orgAn Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goalscs.LG updates on arXiv.orgEvaluating Machine Learning Models for Post-Wildfire Debris-Flow Predictioncs.LG updates on arXiv.orgRectifying Geometric Misalignment: Online Source-Free Adaptation for Class-Imbalanced EEGcs.LG updates on arXiv.orgQEvict: Recoverable Quantized KV Eviction for Attention-Drift-Robust Long-Context Decodingcs.LG updates on arXiv.orgDG-FedReuse: Proxy-Gradient-Gated Cached-Update Reuse with Matched Sparse Uplink Accountingcs.LG updates on arXiv.orgQuantum-Structured World Models (QSWMs) for Predictive Latent Dynamicscs.LG updates on arXiv.orgSpectral Distillation: From Nonlinear Dynamics to Linear State-Space Modelscs.LG updates on arXiv.orgPerturbation Sensitivity at Convergence: A Simple Signal for Identifying Spuriously Correlated Samplescs.LG updates on arXiv.orgIFlowNets: Extending Generative Samplers to Learn Strategies in Incomplete Information GamesQuanta MagazineNeutrinos From Deep Inside Earth Provide a New Picture of the MantleAi2 BlogTutorMoments: Do AI tutors know when to help and when to hold back?cs.LG updates on arXiv.orgMS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classificationcs.LG updates on arXiv.orgWhen Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasterscs.LG updates on arXiv.orgPPDL: LLM-Based Flows as Probabilistic Programscs.LG updates on arXiv.orgDecoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Languagecs.LG updates on arXiv.orgDisentangling 3D Modeling from Spatial Reasoningcs.LG updates on arXiv.orgMarginal Matching Does Not License Factorized Sampling: Auditing Conditional Style Leakage in Factorized Generative Modelscs.LG updates on arXiv.orgPRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesiscs.LG updates on arXiv.orgBeyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuningcs.LG updates on arXiv.orgBeyond Rotations: AuroOFT for Expressive Quantized Orthogonal Fine-Tuningcs.LG updates on arXiv.orgAn Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goalscs.LG updates on arXiv.orgEvaluating Machine Learning Models for Post-Wildfire Debris-Flow Predictioncs.LG updates on arXiv.orgRectifying Geometric Misalignment: Online Source-Free Adaptation for Class-Imbalanced EEGcs.LG updates on arXiv.orgQEvict: Recoverable Quantized KV Eviction for Attention-Drift-Robust Long-Context Decodingcs.LG updates on arXiv.orgDG-FedReuse: Proxy-Gradient-Gated Cached-Update Reuse with Matched Sparse Uplink Accountingcs.LG updates on arXiv.orgQuantum-Structured World Models (QSWMs) for Predictive Latent Dynamicscs.LG updates on arXiv.orgSpectral Distillation: From Nonlinear Dynamics to Linear State-Space Modelscs.LG updates on arXiv.orgPerturbation Sensitivity at Convergence: A Simple Signal for Identifying Spuriously Correlated Samplescs.LG updates on arXiv.orgIFlowNets: Extending Generative Samplers to Learn Strategies in Incomplete Information Games

By Source

Feeds organized so you can skim by site.

Density Sort
QM
Quanta Magazine
21h ago · 5 items
AB
Ai2 Blog
1d ago · 20 items
TutorMoments: Do AI tutors know when to help and when to hold back? 1d ago TutorMoments is an open, replay-based evaluation framework that tests whether AI tutors can recognize when to support a student and when to hold back and encourage deeper reasoning. Ai2 expands collaboration with Hugging Face to accelerate open science 2d ago Ai2 is expanding its partnership with Hugging Face to give its growing portfolio of fully open models, datasets, benchmarks, and applications the storage, bandwidth, and integrations needed to reach more researchers and developers. Tracing distinctive language in AI-written text 8d ago Stony Brook researchers used our infini-gram engine to trace distinctive phrases in AI-generated writing back to existing sources, finding that top-selling self-published books on Amazon with substantial detected AI text overlap more heavil... The OlmoEarth Platform: Geospatial inference at planetary scale 11d ago How we built the OlmoEarth Platform to fine-tune geospatial models and run continent-scale satellite inference while managing massive data pipelines, distributed compute, and automatically recovering from failures at scale. Who gets to understand AI? 15d ago Why fully open models and research artifacts are essential to independent scrutiny, broader participation, and continued U.S. scientific leadership in AI. What building Shippy taught us about building agents 26d ago Building Shippy taught us that reliable agents depend less on the model itself than on deterministic tools, explicit guardrails, isolated infrastructure, and evaluations grounded in real-world workflows and live data. MolmoAct 2 shows what open models can unlock for robotics 31d ago Robotics engineer Binh Pham used MolmoAct 2 to build a voice-controlled robot that won South Park Commons’ embodied AI hackathon. Modular LLMs at scale: how FlexOlmo is helping to pool national expertise without pooling sensitive data 37d ago Danish Foundation Models is using FlexOlmo as the basis for FlexMoRE, a more efficient modular LLM architecture that lets institutions contribute specialized experts trained on sensitive or proprietary data without sharing that data—and run... Which tokens does a hybrid model predict better? 44d ago New token-level analyses of Olmo 3 and Olmo Hybrid show that hybrid models predict meaning-bearing, context-dependent tokens better than transformers, while transformers retain an edge on verbatim copying. How Domyn and AISquared built on Ai2's open releases 51d ago Domyn and AISquared show how Ai2’s open releases are helping AI labs build models for regulated industries, where transparency, provenance, licensing, and control are essential for customer trust and compliance.
20 loaded
CL
cs.LG updates on arXiv.org
1d ago · 1320 items
MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification 1d ago Abstract page for arXiv paper 2608.05196: MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters 1d ago Abstract page for arXiv paper 2608.05207: When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters PPDL: LLM-Based Flows as Probabilistic Programs 1d ago Abstract page for arXiv paper 2608.05234: PPDL: LLM-Based Flows as Probabilistic Programs Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language 1d ago Abstract page for arXiv paper 2608.05238: Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language Disentangling 3D Modeling from Spatial Reasoning 1d ago Abstract page for arXiv paper 2608.05242: Disentangling 3D Modeling from Spatial Reasoning Marginal Matching Does Not License Factorized Sampling: Auditing Conditional Style Leakage in Factorized Generative Models 1d ago Abstract page for arXiv paper 2608.05243: Marginal Matching Does Not License Factorized Sampling: Auditing Conditional Style Leakage in Factorized Generative Models PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis 1d ago Abstract page for arXiv paper 2608.05249: PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning 1d ago Abstract page for arXiv paper 2608.05250: Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning Beyond Rotations: AuroOFT for Expressive Quantized Orthogonal Fine-Tuning 1d ago Abstract page for arXiv paper 2608.05253: Beyond Rotations: AuroOFT for Expressive Quantized Orthogonal Fine-Tuning An Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goals 1d ago Abstract page for arXiv paper 2608.05255: An Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goals
1320 loaded
SM
stat.ML updates on arXiv.org
1d ago · 1351 items
Quality Diversity for Reliable Data Driven Time-Use Optimization 1d ago Abstract page for arXiv paper 2608.05230: Quality Diversity for Reliable Data Driven Time-Use Optimization A Unified Causal Inference Framework for the Desirability of Outcome Ranking Paradigm in Benefit-Risk Evaluation 1d ago Abstract page for arXiv paper 2608.05244: A Unified Causal Inference Framework for the Desirability of Outcome Ranking Paradigm in Benefit-Risk Evaluation Deep Generalised Mixed Models: a Novel Neural Network Structure for Analysing Hierarchical Data 1d ago Abstract page for arXiv paper 2608.05930: Deep Generalised Mixed Models: a Novel Neural Network Structure for Analysing Hierarchical Data Handling Missing Data in Probabilistic Regression Trees 1d ago Abstract page for arXiv paper 2608.06195: Handling Missing Data in Probabilistic Regression Trees Beyond Marginal Validity: Finite-Sample Guarantees for Localized Conformal Prediction 1d ago Abstract page for arXiv paper 2608.06206: Beyond Marginal Validity: Finite-Sample Guarantees for Localized Conformal Prediction Minimax Optimal Early-Stopped Gradient Descent for Gaussian Mixture Classification 1d ago Abstract page for arXiv paper 2608.06250: Minimax Optimal Early-Stopped Gradient Descent for Gaussian Mixture Classification Stochastic Dynamics on Persistence Diagram Space via Reinforcement Learning 1d ago Abstract page for arXiv paper 2608.06276: Stochastic Dynamics on Persistence Diagram Space via Reinforcement Learning Optimal Rates for Learning with Monotone Adversaries 1d ago Abstract page for arXiv paper 2608.06337: Optimal Rates for Learning with Monotone Adversaries Scalable estimation of VARMA models 1d ago Abstract page for arXiv paper 2608.06340: Scalable estimation of VARMA models FlowAdam: Implicit Regularization via Geometry-Aware Soft Momentum Injection 1d ago Abstract page for arXiv paper 2604.06652: FlowAdam: Implicit Regularization via Geometry-Aware Soft Momentum Injection
1351 loaded
AI-based augmentation of oncology clinical trials 1d ago Oncology clinical trials are often characterized by slow accrual, high failure rates and limited generalizability, reflecting both biological complexity and operational inefficiencies. Advances in artificial intelligence (AI) — enabled by l... AI agents are checking the scientific literature — and spotting decades-old errors 2d ago The technology is proving adept at finding faults in decades-old papers and reference databases. Inference of tumor spatial habitats 2d ago Nature Methods - Inference of tumor spatial habitats The Virtual Tissues foundation model resolves spatial proteomics across scales 3d ago Spatial proteomics technologies have transformed our understanding of complex tissue architecture in cancer but present unique challenges for computational analysis1. Each study uses a different marker panel and protocol, and most methods a... Divergent impacts of explainable AI for dermatological diagnosis on clinicians versus lay people 4d ago Artificial intelligence (AI) is increasingly permeating healthcare, from serving as a physician assistant to powering consumer applications. The opacity of AI algorithms makes the ability of humans to interact with AI algorithms challenging... Privacy risks from medical AI tools are not shared equally 4d ago Privacy attacks can reveal whether someone’s medical data was used to train an AI model. People who differ from the majority are the most vulnerable to such attacks. A foundation model for sleep-based risk stratification and clinical outcomes 5d ago Clinical sleep studies capture multiple physiologic signals, yet interpretation is often reduced to single summary measures of limited prognostic value, such as the apnea–hypopnea index. We present a foundation model that learns rich repres... Automatic report-based assessment of radiology-pathology concordance in surgical patients using BERT and DPCNN 5d ago Assessing radiology-pathology concordance is important for retrospective audit, educational feedback, and quality assurance in radiology practice. However, automated concordance assessment remains challenging because of imbalanced data and ... Want to get more from AI? Treat every prompt like an experiment 5d ago Taking a scientific approach to artificial-intelligence queries makes every output a result to be checked, says James Dewar. Here are ten tips for doing it right. Dynamic feature pyramid network for real-time gesture recognition 7d ago Effective gesture recognition in Virtual Reality (VR) and Augmented Reality (AR) faces significant challenges from varying hand angles, postures, and complex backgrounds, limiting real-time application potential. This paper proposes the Dyn...
271 loaded
AM
Apple Machine Learning Research
1d ago · 10 items
Scaling Categorical Flow Maps 1d ago Continuous diffusion and flow matching models could represent a powerful alternative to autoregressive approaches for language modelling… Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models 1d ago Large Language Models (LLMs) have achieved state-of-the-art performance on a broad range of Natural Language Processing (NLP) tasks… Arbitrage: Efficient Reasoning via Advantage-Aware Speculation 1d ago Modern Large Language Models achieve impressive reasoning capabilities with long Chain of Thoughts, but they incur substantial computational… Locking Pretrained Weights via Deep Low-Rank Residual Distillation 2d ago The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by… DeepAmbigQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness 2d ago Large language models (LLMs) with integrated search tools show strong promise in open-domain question answering (QA), yet they often… Taming Outlier Tokens in Diffusion Transformers 3d ago We study outlier tokens in Diffusion Transformers (DiTs) for image generation. Prior work has shown that Vision Transformers (ViTs) can… Understanding Alignment in Multimodal LLMs: A Comprehensive Study 5d ago Preference alignment has become a crucial component in enhancing the performance of Large Language Models (LLMs), yet its impact in… Dimensionality Reduction Meets Network Science: Sensemaking on UMAP’s kNN Graph 9d ago While UMAP is widely used for exploring high-dimensional data, typical workflows focus on its lower-dimensional embedding, largely… MoMo: Dial Motion Mode in Robot Manipulation with Spatiotemporal Action Tokenization 9d ago To operate effectively across diverse contexts, robots must not only perform manipulation tasks accurately but also adapt how their actions… Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers 11d ago Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple’s…
GD
Google DeepMind News
1d ago · 20 items
WeatherNext: AI model achieves breakthrough in forecasting cyclones 1d ago WeatherNext enables accurate cyclone forecasts that can give an extra day of warning. Now we are open sourcing the model. Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration 8d ago Gemini Robotics ER 2 is a step change in video understanding, tool orchestration, and multi-robot collaboration for robotic applications. We’re launching Lyria 3.5 in Google Flow Music, with advances across musicality, lyrics, vocals, and creative control 9d ago Our newest music generation model, Lyria 3.5, delivers significant advancements across musicality, lyrics, and vocal quality, empowering you to craft richer tracks. We’r… Gemini Robotics 2 brings whole body intelligence to robots 10d ago From feet to fingertips — we are teaching robots intelligent whole-body control, fine dexterity, and teamwork to complete a broad range of complex tasks. Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission 16d ago Google is committing $40 million in AI tokens and cloud credits to support the DOE’s Genesis Mission and accelerate groundbreaking scientific discovery. Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber 17d ago We’re introducing new Gemini models, including Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber. Introducing Gemini 3.5 Flash Cyber 21d ago Google introduces Gemini 3.5 Flash Cyber to help defenders find, validate, and patch software vulnerabilities quickly and efficiently. Our approach to bioresilience 23d ago Google DeepMind and Isomorphic Labs approach to bioresilience, using AI models to support prevention, detection and response. Empowering India’s next generation of innovators with ATL Saathi 25d ago Atal Innovation Mission launches ATL Saathi, a Gemini powered AI assistant empowering India's educators to nurture the next generation of innovators. Google DeepMind and A24 announce first-of-its-kind research partnership 35d ago Today, Google DeepMind and A24 are announcing a first-of-its-kind partnership focused on research. The collaboration pairs a world-leading research lab with the industry…
20 loaded
AS
Amazon Science homepage
2d ago · 20 items
34 Amazon Research Awards Build on Trainium recipients announced 2d ago Amazon announces 34 recipients of the Build on Trainium program, a $110 million credit initiative supporting AI research at 30 universities including Stanford, UC Berkeley, UIUC, UCLA, CMU, and MIT, with a focus on Responsible AI. How controllers from industrial machinery can coordinate multitask machine learning 8d ago Instead of compromising among parameter updates dictated by different training objectives, ControlG allocates computational capacity to objectives sequentially and dynamically. A new benchmark for evaluating patient-facing health AI agents 9d ago PatientAgentBench generates a synthetic patient health record, a realistic clinical vignette, and a patient agent that converses with the AI system under evaluation, to capture what a patient-facing agent actually has to do. Amazon is investing in the Lean Focused Research Organization 13d ago As AI agents take on higher-stakes decisions, Lean programming language makes it possible to mathematically prove they will behave safely. Amazon and University of Michigan give robots a sense of touch 28d ago Amazon and University of Michigan researchers developed a physics-based tactile simulator that teaches robots dexterous manipulation skills in simulation with a 93% real-world success rate — no fine-tuning required. Capturing token IDs during agentic interactions for better reinforcement learning 29d ago A new Rust proxy called Turnstile sits between the model backend and the agent harness to capture information lost in mere text transcripts. How Amazon tracks carbon intensity across its operations 37d ago Amazon is developing precise, sector-specific approaches to measuring decarbonization progress — starting with emissions per unit shipped. The fuel of the future is already here: Why TRISO matters 44d ago Each TRISO particle is a millimeter-wide containment system — engineered to withstand extreme temperatures and retain fission products for thousands of years. Here's how this advanced fuel technology works and why Amazon is investing in it ... EC2’s formally verified “isolation engine” provides mathematical assurance of virtual-machine isolation 58d ago 330,000 lines of machine-checked proofs in Isabelle/HOL verify that the Nitro Isolation Engine correctly enforces confidentiality, integrity, and memory safety between EC2 virtual machines on Graviton5. Graviton5’s improved design increases speed and energy efficiency — beyond Moore’s law 58d ago Graviton5's four-chiplet architecture, custom die-to-die connectivity, three-nanometer process, and 192 megabytes of L3 cache deliver up to 35% faster performance for web applications and ML inference.
20 loaded
MN
Solving the solvent problem 3d ago MIT researchers are exploring sodium metal batteries as a cheaper, more abundant alternative for fast, scalable energy storage. The main challenge is sodium metal’s high reactivity, but the researchers show how choosing the right electrolyt... The benefits of medical AI assistance vary based on user expertise 4d ago New research found non-experts deferred to AI-based assistance in diagnosing skin cancer, even when it was wrong, while clinicians were more likely to catch AI errors. Alexander Rakhlin named director of the MIT Statistics and Data Science Center 4d ago Alexander ‘Sasha’ Rakhlin, the Distinguished Professor in Data, Systems, and Society, IDSS and Brain and Cognitive Sciences at MIT, has been named the next director of the MIT Statistics and Data Science Center. Daniela Rus receives Bavarian Minister-President's High-Tech Prize 8d ago MIT Professor and CSAIL Director Daniela Rus has received the 2026 High-Tech Prize of the Bavarian Minister-President for projects like self-organizing robot collectives, soft robotics, autonomous mobility, and brain-inspired artificial int... Connecting research to policy on Capitol Hill 8d ago MIT students and postdocs traveled to Washington to meet with U.S. Senate and House of Representatives staffers. Over two days, they met with over 60 offices from 32 states, advocating for policies ranging from NSF funding to AI safety. How a medical database developed at MIT evolved into a global standard of data-sharing 9d ago The visionary PhysioNet platform launched 25 years ago, based on a system developed at MIT in the 1970s. It has become one of the most comprehensive biomedical and clinical data repositories in existence. Working to automate nuclear plant operations 15d ago For nuclear to be considered as a viable clean energy source, it has to be competitively priced and economical to produce. Lauren Fortier, formerly a Naval officer and now a doctoral student in MIT's Department of Nuclear Science and E... MIT projects selected for funding under US Department of Energy’s Genesis Mission 15d ago MIT researchers are set to contribute to the U.S. Department of Energy’s (DOE) Genesis Mission, with 15 collaborative projects among those selected for funding under Genesis Phase I, DOE has announced. Professor Emeritus Dimitri Bertsekas, influential computer scientist and prolific author, dies at 83 16d ago MIT Professor Emeritus Dimitri Bertsekas, an influential researcher in many AI-related fields, a prolific textbook author, and a talented travel photographer, has died at age 83. Following the questions where they lead 21d ago A profile of MIT Assistant Professor Bailey Flanigan explores how she develops complex computational methods for helping democracy thrive.
20 loaded
MN
MIT News - Machine learning
4d ago · 20 items
The benefits of medical AI assistance vary based on user expertise 4d ago New research found non-experts deferred to AI-based assistance in diagnosing skin cancer, even when it was wrong, while clinicians were more likely to catch AI errors. Alexander Rakhlin named director of the MIT Statistics and Data Science Center 4d ago Alexander ‘Sasha’ Rakhlin, the Distinguished Professor in Data, Systems, and Society, IDSS and Brain and Cognitive Sciences at MIT, has been named the next director of the MIT Statistics and Data Science Center. A better way to turn 2D designs into 3D models for rapid prototyping 23d ago “GIFT” is a new system that teaches vision-language generative AI models to produce accurate, computer-aided design (CAD) programs that can be used to simulate and test 3D objects. The method is more accurate than competing techniques, usin... 3 Questions: Neural transparency and the future of AI design 23d ago MIT Assistant Professor Pat Pataranutaporn describes a new interface that lets everyday users glimpse inside an AI's neural network before their chatbot ever says a word. Can AI build a jet engine? JARVIS Challenge tests role of AI copilots in tough-tech engineering 24d ago MIT's JARVIS Challenge (Jet-engine AI Research and Validation Intensive Sprint) is a new academic competition asking MIT students to explore whether AI can compress the design-build-test cycle so engineers can build faster and better. AI agents create virtual playgrounds to help robots get crucial training data 25d ago The “SceneSmith” system developed by MIT CSAIL researchers uses AI agents to generate lifelike scenes of indoor environments like kitchens and hotels to help robots simulate everyday chores. These 3D worlds are more realistic and diverse th... New method aims to keep kids safe from illegal AI-generated content 26d ago Researchers developed an evaluation procedure that tests generative AI models for harmful capabilities without generating outputs. This could enable auditors to identify open-source models that have been adapted to produce illegal content, ... Tiny robot boats build floating structures 29d ago FloatForm, developed at MIT, is a swarm of small aquatic robots that assemble into reconfigurable structures. It could lead to floating infrastructure that builds itself into things like a temporary platform, a market, or a stage. MIT-designed educational factory embraces modern manufacturing 30d ago The FrED factory, a low-cost desktop fiber extrusion device designed and assembled by students in an educational factory at MIT, is changing how manufacturing is taught in Mexico through a collaboration with Tecnológico de Monterrey. Q&A: What is agentic AI today, and what do we want it to be? 38d ago MIT Associate Professor Phillip Isola explains what agentic AI is, how these systems are used, what applications they are best suited for, and what the future may hold for this exploding technology.
20 loaded
MR
Microsoft Research
4d ago · 10 items
Orchard: An open framework for scalable agentic AI 4d ago Orchard is an open-source framework for the research community to train and evaluate AI agents across task types. It reduces complexity while supporting strong performance from smaller models by enabling researchers to reuse the same infras... Echoverse: Deep, evolving environments for computer-use agents 8d ago Computer-use AI agents struggle with multi-step workflows like email and customer support. Echoverse trains agents in realistic environments rather than simply providing more training tasks, helping them improve as the tasks, tests, and env... EvoLib: Turning experience into evolving knowledge 8d ago LLMs do not get smarter just by remembering more. EvoLib turns experience into evolving knowledge, taking reusable skills and insights that help models learn and adapt across tasks long after deployment. Verifying Rust cryptography in SymCrypt, from standards to code 25d ago Cryptographic code supports vital protections in modern computing systems. Learn how a new method helps verify code as developers write it while preserving speed and adaptability as it gets implemented and evolves: Aurora 1.5: Extending open foundation models for weather and Earth-system applications 29d ago Aurora 1.5 adds 22 more variables, hourly temporal resolution, and probabilistic ensemble forecasting to the Aurora foundation model, making it more useful for real-world weather, climate, and energy applications. Flint: A visualization language for the AI era 30d ago Short chart specifications are easy to write, but often produce uninspiring results. Flint is an open-source visualization language that offers a middle path, letting AI agents create expressive charts from compact, human-editable specifica... SkillOpt: Agent skills as trainable parameters 38d ago AI agents often fail because their instructions, or skills, are manually modified with no guarantee of improvement. Learn how SkillOpt turns skill editing into a training process, making agent behavior more reliable without changing model w... Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity 39d ago AI agents can't remember past conversations. They must constantly reload or retrieve context, which grows less efficient as tasks get longer and more complex. Memora solves this with a scalable memory system separating what’s stored from ho... Understanding the brain with AI-driven explanations and experiments 43d ago Researchers introduce generative causal testing, which translates black box models into clear hypotheses and verifies them in the scanner, revealing what specific brain regions respond to in language. Talos: Scaling rare disease diagnosis with automated, iterative genomic reanalysis 44d ago Talos was built to help resolve a major bottleneck in genomic medicine: human review time. The open-source system recovered 90% of in-scope diagnoses while surfacing just 1.3 candidate variants per patient for expert review.
TL
The latest research from Google
8d ago · 20 items
20 loaded
How Open Models Are Driving AI Research 32d ago NVIDIA open models from Nemotron, Cosmos and BioNeMo are fueling the field's biggest research questions at ICML 2026. NVIDIA Research Unlocks Advanced Grasping, Smarter Autonomous Driving and Agent Training at Scale 65d ago New NVIDIA Research breakthroughs show how training at scale — across gripper types, driving scenarios and virtual worlds — creates AI that generalizes to diverse applications. NVIDIA Enables the Next Era Of Physical AI Research With Agent Skills For Autonomous Vehicles, Robotics And Vision AI 65d ago New physical AI agent skills, powered by NVIDIA Cosmos 3, help researchers accelerate data generation, simulation, policy training and evaluation for autonomous system development. NVIDIA Research Advances Robotics From Simulation to the Real World 71d ago Featured at the International Conference on Robotics and Automation, eight new NVIDIA Research papers show how robots trained in simulation are moving into the real world. NVIDIA Launches Earth-2 Family of Open Models — the World’s First Fully Open, Accelerated Set of Models and Tools for AI Weather 193d ago NVIDIA Earth-2 makes weather AI accessible worldwide at every stage — from processing initial observation data to generating 15-day global forecasts or local storm forecasts. At NeurIPS, NVIDIA Advances Open Model Development for Digital and Physical AI 249d ago NVIDIA releases new AI tools for speech, safety and autonomous driving — including NVIDIA DRIVE Alpamayo-R1, the world’s first open industry-scale reasoning vision language action model for mobility — and a new independent benchmark recogni... How Do You Teach an AI Model to Reason? With Humans 345d ago NVIDIA’s data factory team creates the foundation for AI models like Cosmos Reason, which today topped the physical reasoning leaderboard on Hugging Face. NVIDIA Research Shapes Physical AI 361d ago AI and graphics research breakthroughs in neural rendering, 3D generation and world simulation power robotics, autonomous vehicles and content creation. NVIDIA Research Showcases the Future of Robotics at RSS 413d ago At this year’s Robotics: Science and Systems conference, NVIDIA Research is presenting work that advances robot learning across simulation, real-world transfer and decision-making. NVIDIA Scores Consecutive Win for End-to-End Autonomous Driving Grand Challenge at CVPR 423d ago NVIDIA was today named an Autonomous Grand Challenge winner at the Computer Vision and Pattern Recognition (CVPR) conference, held this week in Nashville, Tennessee. The announcement was made at the Embodied Intelligence for Autonomous Syst...
18 loaded
FO
Future of Life Institute
49d ago · 20 items
Should AIs be people too? 49d ago Statement: Anthropic warns of AI self-improvement risks, considers a pause 61d ago FLI President on the White House Executive Order 65d ago Magnificent Humanity – The Pope’s First Encyclical Concerns AI 79d ago White House working group on AI – Statement from FLI’s Anthony Aguirre 94d ago FLI’s President and CEO on Trump’s support for an AI ‘kill switch’ 113d ago FLI CEO’s statement on the attack against Sam Altman’s home 119d ago Prominent Scientists, Faith Leaders, Policymakers and Artists Call for a Prohibition on Superintelligence, as Poll Shows Americans Don’t Want It 133d ago Statement: Head of US Policy on the White House AI legislative recommendations 138d ago Governor DeSantis Directs Florida State Agencies to Partner with Future of Life Institute to Shield Families from AI Harm 151d ago
20 loaded
IN
inFERENCe
163d ago · 15 items
The Future of Software 163d ago The world of software is undergoing a shift not seen since the advent of compilers in the 1970s. Compilers were the original vibe coding: they automatically generate complex machine code that human programmers had to manually write before. ... Deep Learning is Powerful Because It Makes Hard Things Easy - Reflections 10 Years On 188d ago Ten years ago this week, I wrote a post called "Deep Learning is Easy - Learn Something Harder". The post blew up, top spot on HackerNews. Needless to say, it didn't age well. Discrete Diffusion: Continuous-Time Markov Chains 443d ago A tutorial explaining some intuitions behind continuous time Markov chains for machine learners interested in discrete diffusion models. We may finally crack Maths. But should we? 1156d ago Automating mathematical theorem proving has been a long standing goal of artificial intelligence and indeed computer science. It's one of the areas I became very interested in recently. This is because I feel we may have the ingredients nee... Mortal Komputation: On Hinton's argument for superhuman AI. 1165d ago Last week in Cambridge was Hinton bonanza. He visited the university town where he was once an undergraduate in experimental psychology, and gave a series of back-to-back talks, Q&A sessions, interviews, dinners, etc. He was stopped on the ... Autoregressive Models, OOD Prompts and the Interpolation Regime 1226d ago A few years ago I was very much into maximum likelihood-based generative modeling and autoregressive models (see this, this or this). More recently, my focus shifted to characterising inductive biases of gradient-based optimization focussin... We May be Surprised Again: Why I take LLMs seriously. 1234d ago "Deep Learning is Easy, Learn something Harder" - I proclaimed in one of my early and provocative blog posts from 2016. While some observations were fair, that post is now evidence that I clearly underestimated the impact simple techniques ... Implicit Bayesian Inference in Large Language Models 1618d ago This intriguing paper kept me thinking long enough for me to I decide it's time to resurrect my blogging (I started writing this during ICLR review period, and realised it might be a good idea to wait until that's concluded) * Sang Michael ... Eastern European Guide to Writing Reference Letters 1621d ago Excruciating. One phrase I often use to describe what it's like to read reference letters for Eastern European applicants to PhD and Master's programs in Cambridge. Even objectively outstanding students often receive dull, short, factual, a... Causal inference 4: Causal Diagrams, Markov Factorization, Structural Equation Models 1884d ago This post is written with my PhD student and now guest author Patrik Reizinger [https://twitter.com/rpatrik96] and is part 4 of a series of posts on causal inference: * Part 1: Intro to causal inference and do-calculus [https://www.inferenc...
15 loaded
TG
The Gradient
170d ago · 15 items
15 loaded
VI
VITALab
214d ago · 10 items
Towards Brain MRI Foundation Models for the Clinic: Findings from the FOMO25 Challenge 214d ago 1. Motivation Brain Latent Progression Individual-based spatiotemporal disease progression on 3D Brain MRIs via latent diffusion 345d ago This article aims at reviewing a Alzheimer’s spatiotemporal disease progression predictive model called Brain Latent Progression (BrLP). All in all, this is ... A Survey of popular LLM Evaluation Metrics 353d ago Large Language Models (LLMs) are increasingly applied to critical domains such as medical report generation, where accuracy and trust are essential. Evaluati... Open-Source Large Language Models in Radiology: A Review and Tutorial for Practical Research and Clinical Deployment 362d ago Open-Source Large Language Models in Radiology MemSAM: Taming Segment Anything Model for Echocardiography Video Segmentation 431d ago MemSAM Simplifying Deep Temporal Difference Learning 488d ago tl;dr The authors propose PQN, a simplified deep online Q-Learning that uses very small replay buffers. Normalization and parallelized sampling from vectoriz... EchoPrime: Multi-Video View-Informed Vision-Language Model for Comprehensive Echocardiography Interpretation 502d ago Objective EchoPrime is a foundation model designed for comprehensive echocardiographic interpretation. Unlike previous models that use single views or static... DeepSeek-V3 Technical Report 543d ago DeepSeek-V3 Variational Autoencoders for Generating Synthetic Tractography-Based Bundle Templates in a Low-Data Setting 571d ago Highlights Implicit neural representations 599d ago Implicit neural networks

No matching sources found.