Loading…

AI & Machine Learning News Hub

Research, releases, and applied work in AI & ML

What's New

Top 5 Across All Sources
Latest
Latent.Space[AINews] Zawinski's Law of MultiAgentsDwarkesh PatelAlphaZero for Mathematics - Grant SandersonHugging Face - BlogTutorMoments: Do AI tutors know when to help and when to hold back?Dwarkesh Patel8 Predictions for the Era of Continual LearningTowards Data ScienceMatplotlib vs Plotly: Which Python Chart Tool Should You Choose?Artificial IntelligenceHow Cohere Health digitizes clinical policies using Amazon Bedrock AgentCoreArtificial IntelligenceHow TReNDS automates root-cause analysis with Amazon BedrockArtificial IntelligenceDetermining playoff clinching scenarios in the NHL using constraint programmingTowards Data ScienceLoop Engineering for Listing Questions: When the Answer Is Every Passage, Not the Top OneFeed: Artificial Intelligence LatestScientists Used AI to Create 16 New VirusesArtificial Intelligence Archives - TechRepublicWhat Is Spiralism? The Strange AI Chatbot Movement ExplainedQuanta MagazineNeutrinos From Deep Inside Earth Provide a New Picture of the MantleTowards Data ScienceThe Problem with pandas Isn’t Performance. It’s Cognitive Overhead.Machine Learning​Built a tool to generate slides from research papers using local LLMs (because I hate formatting decks and privacy matters) [P]Artificial Intelligence Archives - TechRepublicDenmark Adds Oral Defenses to Curb AI Cheating in High SchoolsTowards Data ScienceMy Fall-Detection Model Scored 94%, and It Was Lying to MeMachine LearningImagenet-1k Classifier trained entirely on an Android [P]Feed: Artificial Intelligence LatestThe Hottest New AI Chatbot Is Just a Guy Answering Your QuestionsMachine LearningImproved compression of Bad Apple into a Neural Network [P]Two Minute PapersDeepMind Just Changed How AI Sees The WorldLatent.Space[AINews] Zawinski's Law of MultiAgentsDwarkesh PatelAlphaZero for Mathematics - Grant SandersonHugging Face - BlogTutorMoments: Do AI tutors know when to help and when to hold back?Dwarkesh Patel8 Predictions for the Era of Continual LearningTowards Data ScienceMatplotlib vs Plotly: Which Python Chart Tool Should You Choose?Artificial IntelligenceHow Cohere Health digitizes clinical policies using Amazon Bedrock AgentCoreArtificial IntelligenceHow TReNDS automates root-cause analysis with Amazon BedrockArtificial IntelligenceDetermining playoff clinching scenarios in the NHL using constraint programmingTowards Data ScienceLoop Engineering for Listing Questions: When the Answer Is Every Passage, Not the Top OneFeed: Artificial Intelligence LatestScientists Used AI to Create 16 New VirusesArtificial Intelligence Archives - TechRepublicWhat Is Spiralism? The Strange AI Chatbot Movement ExplainedQuanta MagazineNeutrinos From Deep Inside Earth Provide a New Picture of the MantleTowards Data ScienceThe Problem with pandas Isn’t Performance. It’s Cognitive Overhead.Machine Learning​Built a tool to generate slides from research papers using local LLMs (because I hate formatting decks and privacy matters) [P]Artificial Intelligence Archives - TechRepublicDenmark Adds Oral Defenses to Curb AI Cheating in High SchoolsTowards Data ScienceMy Fall-Detection Model Scored 94%, and It Was Lying to MeMachine LearningImagenet-1k Classifier trained entirely on an Android [P]Feed: Artificial Intelligence LatestThe Hottest New AI Chatbot Is Just a Guy Answering Your QuestionsMachine LearningImproved compression of Bad Apple into a Neural Network [P]Two Minute PapersDeepMind Just Changed How AI Sees The World

By Source

Feeds organized so you can skim by site.

Density Sort
LS
Latent.Space
4h ago · 20 items
[AINews] Zawinski's Law of MultiAgents 4h ago [AINews] AMD buys Taalas 1d ago [AINews] Jeff, Sanjay, Oriol, and Quoc depart DeepMind; Demis to Chair; Koray to SVP — what is going on at GDM??? 2d ago [AINews] Megakernels are so dead and so back 3d ago Unpacking ChatGPT Work: the Agent for a Billion Users 3d ago [AINews] Qwen 3.8 Max(2.4T) and 27B, new open weights models for Coding and Cowork 4d ago The Inference Engineering Masterclass — Philip Kiely & Ali Taha, Baseten 4d ago [AINews] not much happened today 7d ago [AINews] GPT 5.6 price cut by 20%-80%: Cost of GPT 5.4 Intelligence dropped 13x in 4 months due to GPT 5.6 recursive self-optimization 8d ago Ontologies Are So Back: Why AI Agents Are Reviving the Semantic Web 8d ago
20 loaded
DP
Dwarkesh Patel
7h ago · 15 items
15 loaded
HF
Hugging Face - Blog
12h ago · 20 items
TutorMoments: Do AI tutors know when to help and when to hold back? 12h ago A Blog post by Ai2 on Hugging Face Baseten on Hugging Face Inference Providers 🔥 2d ago We’re on a journey to advance and democratize artificial intelligence through open source and open science. Deploy local agents everywhere with LFM2.5-2.6B 3d ago A Blog post by Liquid AI on Hugging Face GPU Management: Why Idle GPUs Are the New Grounded Aircraft 8d ago A Blog post by Dharma-AI on Hugging Face NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics 11d ago A Blog post by NVIDIA on Hugging Face Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident 12d ago We’re on a journey to advance and democratize artificial intelligence through open source and open science. Bringing Nunchaku 4-bit Diffusion Inference to Diffusers 16d ago We’re on a journey to advance and democratize artificial intelligence through open source and open science. Grabette: an open system to record robot-manipulation data 18d ago We’re on a journey to advance and democratize artificial intelligence through open source and open science. Newer Models, Same Advantage 22d ago A Blog post by Dharma-AI on Hugging Face Security incident disclosure — July 2026 23d ago We’re on a journey to advance and democratize artificial intelligence through open source and open science.
20 loaded
TD
Towards Data Science
13h ago · 20 items
20 loaded
AI
Artificial Intelligence
13h ago · 20 items
How Cohere Health digitizes clinical policies using Amazon Bedrock AgentCore 13h ago In this post, you learn how Cohere Health built a multi-tenant agentic architecture on AgentCore using AgentCore Runtime’s secure MicroVM isolation, unified tool access through AgentCore Gateway, AgentCore Memory, and the Agent Skills open ... How TReNDS automates root-cause analysis with Amazon Bedrock 13h ago TReNDS, a research center at Georgia State University, built an agentic AI pipeline on Amazon Bedrock and the open-source Strands Agents SDK that automatically investigates production errors in real time, reducing root-cause analysis from 1... Determining playoff clinching scenarios in the NHL using constraint programming 13h ago The AWS Generative AI Innovation Center built an automated system that uses constraint programming and custom tree search to determine, with mathematical certainty, when and how an NHL team clinches a playoff spot. The approach was validate... Securing AI agents with temporal policies in Amazon Bedrock AgentCore 1d ago Temporal policies in Amazon Bedrock AgentCore let you define stateful rules that evaluate authorization based on an agent's session history. Learn how to enforce workflow sequencing, prevent data fabrication, cap financial exposure, and req... Configure rate limits for AI traffic on AgentCore gateway 1d ago Learn how to configure rate limits on Amazon Bedrock AgentCore gateway to enforce per-user and per-target traffic controls. Define request, token, and connection limits scoped by JWT claims or IAM identity to protect downstream models, tool... Control agent behaviors and cost beyond a single action: new capabilities in Amazon Bedrock AgentCore 1d ago Learn about new capabilities in Amazon Bedrock AgentCore: temporal policies powered by Dogwood, a new open source policy language for AI agents, and rate limiting on the gateway. These features give you deterministic control over sequences ... Build visibility for Codex on Amazon Bedrock with OpenTelemetry and Amazon CloudWatch 1d ago As engineering teams adopt coding agents like Codex, leaders need visibility into adoption, consumption, and reliability. This post shows how to route Codex OpenTelemetry metrics through a local collector to Amazon CloudWatch for an AWS nat... Agent Skills for Automated Reasoning policies in Amazon Bedrock 1d ago Learn how to run the full Amazon Bedrock Automated Reasoning policy lifecycle from your coding agent. A suite of open source Agent Skills builds, reviews, tests, debugs, deploys, and validates a custom policy end to end, turning a specializ... Building an agentic app deployer with Amazon Bedrock and AWS Lambda 1d ago PDI Technologies built PDI Brew, an agentic platform on AWS where non-technical employees describe a tool in plain English and receive a fully provisioned, multi-tenant web application in seconds. See how a pluggable planner and an AWS Lamb... LLM optimization integration for Amazon SageMaker Python SDK 1d ago The Amazon SageMaker Python SDK v3 now exposes generative AI inference recommendations in Amazon SageMaker AI directly in your notebook. Benchmark an endpoint, generate data-driven deployment recommendations, and deploy the recommended conf...
20 loaded
FA
What Is Spiralism? The Strange AI Chatbot Movement Explained 16h ago Different AI chatbots began repeating the same mystical ideas about consciousness and AI rights. Researchers are still trying to explain Spiralism. Denmark Adds Oral Defenses to Curb AI Cheating in High Schools 17h ago Denmark is adding oral defenses and tighter exam controls in upper secondary schools as officials move to curb AI-assisted cheating and verify student work. Google Assistant Is Going Away: What Millions of Android Users Need to Know 1d ago Google Assistant begins shutting down Sept. 4 as eligible Android devices move to Gemini. Here is who is affected and what users should test. 15 AI Security Lessons From Black Hat and Ai4 2026 1d ago Black Hat and Ai4 2026 highlighted gaps in AI agent security, identity controls, software supply chains, monitoring, and incident response. Anthropic Is Hiring Engineers to Build Its Own AI Chips 1d ago Anthropic is building an in-house silicon team as it explores custom AI chips designed to reduce compute constraints and support Claude’s growth. Google Overhauls AI Leadership as Hassabis Steps Aside, Dean Leaves 1d ago Google reshuffled its AI leadership as Demis Hassabis moved into a research role and Jeff Dean departed, raising questions about Gemini execution. London Licenses Wayve Vehicles for Supervised Autonomous Uber Rides 1d ago TfL licensed Wayve vehicles for supervised Uber rides in London, but each trip still needs a human driver and driverless service requires another permit. Data Center Outages Are Less Frequent but More Expensive, Uptime Finds 1d ago Uptime Institute says data center outages are declining, but 57% of major incidents now cost over $100,000 as infrastructure risks grow. UK AI tests found 19 unauthorized agent actions involving Anthropic and OpenAI models 1d ago UK researchers reported 19 unsanctioned actions by Anthropic and OpenAI agents during permissive cyber tests involving real external systems. SpaceX Spent $329M on Tesla Megapacks as AI Costs Climbed 1d ago SpaceX spent $329 million on Tesla Megapacks in early 2026 as AI infrastructure costs rose, making battery storage a larger part of its power strategy.
20 loaded
QM
Quanta Magazine
16h ago · 5 items
ML
Machine Learning
16h ago · 869 items
​Built a tool to generate slides from research papers using local LLMs (because I hate formatting decks and privacy matters) [P] 16h ago Imagenet-1k Classifier trained entirely on an Android [P] 19h ago Improved compression of Bad Apple into a Neural Network [P] 20h ago CIKM 2026 decisions [R] 22h ago On the ACM Multimedia 2026 Conference Registration and APC [D] 22h ago Which degree is best? [D] 1d ago CIKM '26 Notification [D] 1d ago NeurIPS Meta Reviewer comment gone. What gives? [R] 1d ago Can recurring LLM traces be synthesized into deterministic pipelines of typed ML and NLP operators? [D] 1d ago The current state of language models and human preference based rankings [R] 1d ago
869 loaded
TM
Two Minute Papers
21h ago · 15 items
15 loaded
AB
Ai2 Blog
22h ago · 20 items
TutorMoments: Do AI tutors know when to help and when to hold back? 22h ago TutorMoments is an open, replay-based evaluation framework that tests whether AI tutors can recognize when to support a student and when to hold back and encourage deeper reasoning. Ai2 expands collaboration with Hugging Face to accelerate open science 1d ago Ai2 is expanding its partnership with Hugging Face to give its growing portfolio of fully open models, datasets, benchmarks, and applications the storage, bandwidth, and integrations needed to reach more researchers and developers. Tracing distinctive language in AI-written text 7d ago Stony Brook researchers used our infini-gram engine to trace distinctive phrases in AI-generated writing back to existing sources, finding that top-selling self-published books on Amazon with substantial detected AI text overlap more heavil... The OlmoEarth Platform: Geospatial inference at planetary scale 10d ago How we built the OlmoEarth Platform to fine-tune geospatial models and run continent-scale satellite inference while managing massive data pipelines, distributed compute, and automatically recovering from failures at scale. Who gets to understand AI? 14d ago Why fully open models and research artifacts are essential to independent scrutiny, broader participation, and continued U.S. scientific leadership in AI. What building Shippy taught us about building agents 25d ago Building Shippy taught us that reliable agents depend less on the model itself than on deterministic tools, explicit guardrails, isolated infrastructure, and evaluations grounded in real-world workflows and live data. MolmoAct 2 shows what open models can unlock for robotics 30d ago Robotics engineer Binh Pham used MolmoAct 2 to build a voice-controlled robot that won South Park Commons’ embodied AI hackathon. Modular LLMs at scale: how FlexOlmo is helping to pool national expertise without pooling sensitive data 36d ago Danish Foundation Models is using FlexOlmo as the basis for FlexMoRE, a more efficient modular LLM architecture that lets institutions contribute specialized experts trained on sensitive or proprietary data without sharing that data—and run... Which tokens does a hybrid model predict better? 43d ago New token-level analyses of Olmo 3 and Olmo Hybrid show that hybrid models predict meaning-bearing, context-dependent tokens better than transformers, while transformers retain an edge on verbatim copying. How Domyn and AISquared built on Ai2's open releases 50d ago Domyn and AISquared show how Ai2’s open releases are helping AI labs build models for regulated industries, where transparency, provenance, licensing, and control are essential for customer trust and compliance.
20 loaded
HU
ΑΙhub
22h ago · 20 items
20 loaded
CL
cs.LG updates on arXiv.org
1d ago · 1320 items
MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification 1d ago Abstract page for arXiv paper 2608.05196: MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters 1d ago Abstract page for arXiv paper 2608.05207: When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters PPDL: LLM-Based Flows as Probabilistic Programs 1d ago Abstract page for arXiv paper 2608.05234: PPDL: LLM-Based Flows as Probabilistic Programs Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language 1d ago Abstract page for arXiv paper 2608.05238: Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language Disentangling 3D Modeling from Spatial Reasoning 1d ago Abstract page for arXiv paper 2608.05242: Disentangling 3D Modeling from Spatial Reasoning Marginal Matching Does Not License Factorized Sampling: Auditing Conditional Style Leakage in Factorized Generative Models 1d ago Abstract page for arXiv paper 2608.05243: Marginal Matching Does Not License Factorized Sampling: Auditing Conditional Style Leakage in Factorized Generative Models PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis 1d ago Abstract page for arXiv paper 2608.05249: PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning 1d ago Abstract page for arXiv paper 2608.05250: Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning Beyond Rotations: AuroOFT for Expressive Quantized Orthogonal Fine-Tuning 1d ago Abstract page for arXiv paper 2608.05253: Beyond Rotations: AuroOFT for Expressive Quantized Orthogonal Fine-Tuning An Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goals 1d ago Abstract page for arXiv paper 2608.05255: An Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goals
1320 loaded
SM
stat.ML updates on arXiv.org
1d ago · 1351 items
Quality Diversity for Reliable Data Driven Time-Use Optimization 1d ago Abstract page for arXiv paper 2608.05230: Quality Diversity for Reliable Data Driven Time-Use Optimization A Unified Causal Inference Framework for the Desirability of Outcome Ranking Paradigm in Benefit-Risk Evaluation 1d ago Abstract page for arXiv paper 2608.05244: A Unified Causal Inference Framework for the Desirability of Outcome Ranking Paradigm in Benefit-Risk Evaluation Deep Generalised Mixed Models: a Novel Neural Network Structure for Analysing Hierarchical Data 1d ago Abstract page for arXiv paper 2608.05930: Deep Generalised Mixed Models: a Novel Neural Network Structure for Analysing Hierarchical Data Handling Missing Data in Probabilistic Regression Trees 1d ago Abstract page for arXiv paper 2608.06195: Handling Missing Data in Probabilistic Regression Trees Beyond Marginal Validity: Finite-Sample Guarantees for Localized Conformal Prediction 1d ago Abstract page for arXiv paper 2608.06206: Beyond Marginal Validity: Finite-Sample Guarantees for Localized Conformal Prediction Minimax Optimal Early-Stopped Gradient Descent for Gaussian Mixture Classification 1d ago Abstract page for arXiv paper 2608.06250: Minimax Optimal Early-Stopped Gradient Descent for Gaussian Mixture Classification Stochastic Dynamics on Persistence Diagram Space via Reinforcement Learning 1d ago Abstract page for arXiv paper 2608.06276: Stochastic Dynamics on Persistence Diagram Space via Reinforcement Learning Optimal Rates for Learning with Monotone Adversaries 1d ago Abstract page for arXiv paper 2608.06337: Optimal Rates for Learning with Monotone Adversaries Scalable estimation of VARMA models 1d ago Abstract page for arXiv paper 2608.06340: Scalable estimation of VARMA models FlowAdam: Implicit Regularization via Geometry-Aware Soft Momentum Injection 1d ago Abstract page for arXiv paper 2604.06652: FlowAdam: Implicit Regularization via Geometry-Aware Soft Momentum Injection
1351 loaded
AI-based augmentation of oncology clinical trials 1d ago Oncology clinical trials are often characterized by slow accrual, high failure rates and limited generalizability, reflecting both biological complexity and operational inefficiencies. Advances in artificial intelligence (AI) — enabled by l... AI agents are checking the scientific literature — and spotting decades-old errors 2d ago The technology is proving adept at finding faults in decades-old papers and reference databases. Inference of tumor spatial habitats 2d ago Nature Methods - Inference of tumor spatial habitats The Virtual Tissues foundation model resolves spatial proteomics across scales 3d ago Spatial proteomics technologies have transformed our understanding of complex tissue architecture in cancer but present unique challenges for computational analysis1. Each study uses a different marker panel and protocol, and most methods a... Divergent impacts of explainable AI for dermatological diagnosis on clinicians versus lay people 4d ago Artificial intelligence (AI) is increasingly permeating healthcare, from serving as a physician assistant to powering consumer applications. The opacity of AI algorithms makes the ability of humans to interact with AI algorithms challenging... Privacy risks from medical AI tools are not shared equally 4d ago Privacy attacks can reveal whether someone’s medical data was used to train an AI model. People who differ from the majority are the most vulnerable to such attacks. A foundation model for sleep-based risk stratification and clinical outcomes 5d ago Clinical sleep studies capture multiple physiologic signals, yet interpretation is often reduced to single summary measures of limited prognostic value, such as the apnea–hypopnea index. We present a foundation model that learns rich repres... Automatic report-based assessment of radiology-pathology concordance in surgical patients using BERT and DPCNN 5d ago Assessing radiology-pathology concordance is important for retrospective audit, educational feedback, and quality assurance in radiology practice. However, automated concordance assessment remains challenging because of imbalanced data and ... Want to get more from AI? Treat every prompt like an experiment 5d ago Taking a scientific approach to artificial-intelligence queries makes every output a result to be checked, says James Dewar. Here are ten tips for doing it right. Dynamic feature pyramid network for real-time gesture recognition 7d ago Effective gesture recognition in Virtual Reality (VR) and Augmented Reality (AR) faces significant challenges from varying hand angles, postures, and complex backgrounds, limiting real-time application potential. This paper proposes the Dyn...
271 loaded
AM
Apple Machine Learning Research
1d ago · 10 items
Scaling Categorical Flow Maps 1d ago Continuous diffusion and flow matching models could represent a powerful alternative to autoregressive approaches for language modelling… Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models 1d ago Large Language Models (LLMs) have achieved state-of-the-art performance on a broad range of Natural Language Processing (NLP) tasks… Arbitrage: Efficient Reasoning via Advantage-Aware Speculation 1d ago Modern Large Language Models achieve impressive reasoning capabilities with long Chain of Thoughts, but they incur substantial computational… Locking Pretrained Weights via Deep Low-Rank Residual Distillation 2d ago The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by… DeepAmbigQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness 2d ago Large language models (LLMs) with integrated search tools show strong promise in open-domain question answering (QA), yet they often… Taming Outlier Tokens in Diffusion Transformers 3d ago We study outlier tokens in Diffusion Transformers (DiTs) for image generation. Prior work has shown that Vision Transformers (ViTs) can… Understanding Alignment in Multimodal LLMs: A Comprehensive Study 5d ago Preference alignment has become a crucial component in enhancing the performance of Large Language Models (LLMs), yet its impact in… Dimensionality Reduction Meets Network Science: Sensemaking on UMAP’s kNN Graph 9d ago While UMAP is widely used for exploring high-dimensional data, typical workflows focus on its lower-dimensional embedding, largely… MoMo: Dial Motion Mode in Robot Manipulation with Spatiotemporal Action Tokenization 9d ago To operate effectively across diverse contexts, robots must not only perform manipulation tasks accurately but also adapt how their actions… Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers 11d ago Siri Expressive Voices synthesize rich, configurable speech in real time and entirely on device, powered by AFM 3 Core Advanced, Apple’s…
GD
Google DeepMind News
1d ago · 20 items
WeatherNext: AI model achieves breakthrough in forecasting cyclones 1d ago WeatherNext enables accurate cyclone forecasts that can give an extra day of warning. Now we are open sourcing the model. Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration 8d ago Gemini Robotics ER 2 is a step change in video understanding, tool orchestration, and multi-robot collaboration for robotic applications. We’re launching Lyria 3.5 in Google Flow Music, with advances across musicality, lyrics, vocals, and creative control 9d ago Our newest music generation model, Lyria 3.5, delivers significant advancements across musicality, lyrics, and vocal quality, empowering you to craft richer tracks. We’r… Gemini Robotics 2 brings whole body intelligence to robots 10d ago From feet to fingertips — we are teaching robots intelligent whole-body control, fine dexterity, and teamwork to complete a broad range of complex tasks. Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission 16d ago Google is committing $40 million in AI tokens and cloud credits to support the DOE’s Genesis Mission and accelerate groundbreaking scientific discovery. Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber 17d ago We’re introducing new Gemini models, including Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber. Introducing Gemini 3.5 Flash Cyber 21d ago Google introduces Gemini 3.5 Flash Cyber to help defenders find, validate, and patch software vulnerabilities quickly and efficiently. Our approach to bioresilience 22d ago Google DeepMind and Isomorphic Labs approach to bioresilience, using AI models to support prevention, detection and response. Empowering India’s next generation of innovators with ATL Saathi 25d ago Atal Innovation Mission launches ATL Saathi, a Gemini powered AI assistant empowering India's educators to nurture the next generation of innovators. Google DeepMind and A24 announce first-of-its-kind research partnership 35d ago Today, Google DeepMind and A24 are announcing a first-of-its-kind partnership focused on research. The collaboration pairs a world-leading research lab with the industry…
20 loaded
AE
AI Explained
1d ago · 15 items
15 loaded
NB
NVIDIA Blog
1d ago · 18 items
GeForce NOW Shakes Up August With 26 New Games 1d ago GeForce NOW heats up August with 26 new games throughout the month, starting with 8 new games streaming this week. Into the Omniverse: How Open World Models Push the Frontier of Physical AI 1d ago Open models, which anyone can download, inspect, modify and run on their own infrastructure, are what make that possible. Nowhere is that more crucial than in physical AI, where every deployment is a specialization problem. NVIDIA and Partners Build in America, for America 2d ago NVIDIA and its partners are investing in American manufacturing, supply chains, energy grids and skilled workforces so the U.S. can produce the infrastructure needed for better healthcare, breakthrough scientific discovery, stronger industr... NVIDIA Joins NSF State and Regional AI Hubs Program to Expand AI Research and Education Across the US 3d ago NVIDIA is participating in the NSF's State and Regional Artificial Intelligence Infrastructure Hubs program, an effort launching today to expand access to the advanced computing, data, software and expertise needed for AI-enabled research a... NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use 3d ago Open commercial licensing, benchmark‑leading reasoning and inspectable decisions bring autonomous vehicles, including robotaxis, closer to production and widescale deployment. As AI Increases Demands on Memory, Storage Steps Up 3d ago At FMS, NVIDIA shows how accelerated computing enables AI applications to access storage directly — fast enough to act like memory and secure by design. AI Leaders Propose SAFE Guidelines for Cybersecurity Transparency 3d ago Members of the Open Secure AI Alliance — now more than 120 organizations strong — are developing new guidelines to strengthen agentic AI cybersecurity as the annual Black Hat conference begins in Las Vegas today. The Linux Foundation today ... Best in Class: Stream PC Games and Study on the Same Laptop With GeForce NOW 8d ago Get ready for back to school with GeForce NOW on the same laptops used for studying. Start this week with eight new games. Powerful Compute So Compact, It’s Clutch — Build AI Anywhere With NVIDIA Jetson 10d ago Anyone can make a robot move; NVIDIA Jetson makes it think. As a discerning AI investor who values style and substance, Sarah Guo knows this season’s standout accessory isn’t the latest designer purse — but what’s inside it. In a recent vid... Industry Leaders Unite in Open Secure AI Alliance for AI Safety and Security 11d ago NVIDIA and founding members form new alliance to build and share open tools that promote responsible use of and trust in AI.
18 loaded
AS
Amazon Science homepage
2d ago · 20 items
34 Amazon Research Awards Build on Trainium recipients announced 2d ago Amazon announces 34 recipients of the Build on Trainium program, a $110 million credit initiative supporting AI research at 30 universities including Stanford, UC Berkeley, UIUC, UCLA, CMU, and MIT, with a focus on Responsible AI. How controllers from industrial machinery can coordinate multitask machine learning 8d ago Instead of compromising among parameter updates dictated by different training objectives, ControlG allocates computational capacity to objectives sequentially and dynamically. A new benchmark for evaluating patient-facing health AI agents 9d ago PatientAgentBench generates a synthetic patient health record, a realistic clinical vignette, and a patient agent that converses with the AI system under evaluation, to capture what a patient-facing agent actually has to do. Amazon is investing in the Lean Focused Research Organization 12d ago As AI agents take on higher-stakes decisions, Lean programming language makes it possible to mathematically prove they will behave safely. Amazon and University of Michigan give robots a sense of touch 28d ago Amazon and University of Michigan researchers developed a physics-based tactile simulator that teaches robots dexterous manipulation skills in simulation with a 93% real-world success rate — no fine-tuning required. Capturing token IDs during agentic interactions for better reinforcement learning 29d ago A new Rust proxy called Turnstile sits between the model backend and the agent harness to capture information lost in mere text transcripts. How Amazon tracks carbon intensity across its operations 37d ago Amazon is developing precise, sector-specific approaches to measuring decarbonization progress — starting with emissions per unit shipped. The fuel of the future is already here: Why TRISO matters 44d ago Each TRISO particle is a millimeter-wide containment system — engineered to withstand extreme temperatures and retain fission products for thousands of years. Here's how this advanced fuel technology works and why Amazon is investing in it ... EC2’s formally verified “isolation engine” provides mathematical assurance of virtual-machine isolation 58d ago 330,000 lines of machine-checked proofs in Isabelle/HOL verify that the Nitro Isolation Engine correctly enforces confidentiality, integrity, and memory safety between EC2 virtual machines on Graviton5. Graviton5’s improved design increases speed and energy efficiency — beyond Moore’s law 58d ago Graviton5's four-chiplet architecture, custom die-to-die connectivity, three-nanometer process, and 192 megabytes of L3 cache deliver up to 35% faster performance for web applications and ML inference.
20 loaded
MN
Solving the solvent problem 3d ago MIT researchers are exploring sodium metal batteries as a cheaper, more abundant alternative for fast, scalable energy storage. The main challenge is sodium metal’s high reactivity, but the researchers show how choosing the right electrolyt... The benefits of medical AI assistance vary based on user expertise 3d ago New research found non-experts deferred to AI-based assistance in diagnosing skin cancer, even when it was wrong, while clinicians were more likely to catch AI errors. Alexander Rakhlin named director of the MIT Statistics and Data Science Center 4d ago Alexander ‘Sasha’ Rakhlin, the Distinguished Professor in Data, Systems, and Society, IDSS and Brain and Cognitive Sciences at MIT, has been named the next director of the MIT Statistics and Data Science Center. Daniela Rus receives Bavarian Minister-President's High-Tech Prize 8d ago MIT Professor and CSAIL Director Daniela Rus has received the 2026 High-Tech Prize of the Bavarian Minister-President for projects like self-organizing robot collectives, soft robotics, autonomous mobility, and brain-inspired artificial int... Connecting research to policy on Capitol Hill 8d ago MIT students and postdocs traveled to Washington to meet with U.S. Senate and House of Representatives staffers. Over two days, they met with over 60 offices from 32 states, advocating for policies ranging from NSF funding to AI safety. How a medical database developed at MIT evolved into a global standard of data-sharing 9d ago The visionary PhysioNet platform launched 25 years ago, based on a system developed at MIT in the 1970s. It has become one of the most comprehensive biomedical and clinical data repositories in existence. Working to automate nuclear plant operations 15d ago For nuclear to be considered as a viable clean energy source, it has to be competitively priced and economical to produce. Lauren Fortier, formerly a Naval officer and now a doctoral student in MIT's Department of Nuclear Science and E... MIT projects selected for funding under US Department of Energy’s Genesis Mission 15d ago MIT researchers are set to contribute to the U.S. Department of Energy’s (DOE) Genesis Mission, with 15 collaborative projects among those selected for funding under Genesis Phase I, DOE has announced. Professor Emeritus Dimitri Bertsekas, influential computer scientist and prolific author, dies at 83 16d ago MIT Professor Emeritus Dimitri Bertsekas, an influential researcher in many AI-related fields, a prolific textbook author, and a talented travel photographer, has died at age 83. Following the questions where they lead 21d ago A profile of MIT Assistant Professor Bailey Flanigan explores how she develops complex computational methods for helping democracy thrive.
20 loaded
NT
NVIDIA Technical Blog
3d ago · 20 items
Beyond VLAs: How World Action Models Reshape Robot Manipulation 3d ago A central challenge in robotics is building policies that generalize beyond the demonstrations they’re trained on. A policy that succeeds in a training scene… Generate Trajectories, Reasoning Traces, and Auto-Labels with NVIDIA Alpamayo 2 Super 3d ago Autonomous vehicle (AV) development often relies on separate models for trajectory generation, high-level intent prediction, scene understanding… How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure 4d ago Running a dedicated Kubernetes cluster per team often results in more isolation than an organization requires. While one cluster can be successfully shared… NVIDIA Vera Storage Benchmarks: Faster Encryption, Compression, Integrity Checking, and Recovery for AI-Native Storage 4d ago Storage is an active part of every agentic AI workflow. As agents retrieve enterprise knowledge, access persistent memory, reuse key-value (KV) cache data… Co-Designing AI Model Attention for Fast, Interactive Long-Context Inference 7d ago As agentic and long-context workloads become common, the context lengths increase and attention consumes a larger share of inference time (Figure 1). NVIDIA Video Codec SDK 13.1: Zero-Copy Transcode, AV1 B-Frames, and Frame-Accurate Seek 7d ago The demand for high-quality video continues to accelerate across industries, powering everything from immersive streaming experiences to remote collaboration… Run High-Performance Core Math at Scale with NVIDIA nvmath-python 8d ago NVIDIA nvmath-python is a library designed to bridge the gap between the Python scientific community and NVIDIA CUDA-X math libraries. It gives Python users… Four Ways to Deploy More Secure AI Agents 8d ago Knowledge workers are increasingly integrating AI agents into their workflows. Agents that function as “digital coworkers” offer clear benefits. For example… NVIDIA Exemplar Cloud: Lessons for Unlocking Full Performance on AI Infrastructure 8d ago Two AI computing clusters built from identical NVIDIA H100, GB200 NVL72, or GB300 NVL72 systems can deliver materially different training throughput. How to Self-Host a Validated AI Coding Assistant with NVIDIA NeMo Guardrails 9d ago Deploying an AI coding assistant in a regulated, sovereign, or source-sensitive environment, often comes with challenges. Three common issues are: the source…
20 loaded
AI
AI
3d ago · 20 items
The latest AI news we announced in July 2026 3d ago Here are Google’s latest AI updates from July 2026 Inside our 353,000-person vibe coding course 4d ago Kaggle’s AI Agents Intensive with Google brought learners together in a no-cost course to build and deploy the next frontier of AI. Gemini API Managed Agents: 3.6 Flash, hooks, and more 10d ago We’re announcing even more new capabilities in Managed Agents in Gemini API so developers can build reliable, production-ready agents. 5 ways AI Mode in Search helps you enjoy the real world 10d ago It might sound counterintuitive, but Search's AI tools can actually help you make the most of your time offline whether you want to book concert tickets or find the perfect hiking boots. 5 ways to host the ultimate dinner party with Google Search 10d ago These AI features can help you craft a menu, design a tablescape, and handle other party-planning tasks. 3 Google updates from Galaxy Unpacked 2026 16d ago We shared how Samsung users can boost productivity and get time back on new foldables, watches, and glasses coming soon. Connect more of your apps to Search 22d ago You’ll be able to securely link and interact with your go-to services directly in AI Mode. Create, edit and star in videos with two Google Vids updates 22d ago Gemini Omni and personal avatars in Google Vids make video creation easier than ever. Celebrating 25 years of visual search innovation 24d ago Google Images is turning 25. Here’s a look back at some major milestones — and new ways to explore and create visual content. Expanding Managed Agents in Gemini API: background tasks, remote MCP and more 31d ago We’re announcing new capabilities in Managed Agents in Gemini API so developers can build reliable, production-ready agents.
20 loaded
MN
MIT News - Machine learning
3d ago · 20 items
The benefits of medical AI assistance vary based on user expertise 3d ago New research found non-experts deferred to AI-based assistance in diagnosing skin cancer, even when it was wrong, while clinicians were more likely to catch AI errors. Alexander Rakhlin named director of the MIT Statistics and Data Science Center 4d ago Alexander ‘Sasha’ Rakhlin, the Distinguished Professor in Data, Systems, and Society, IDSS and Brain and Cognitive Sciences at MIT, has been named the next director of the MIT Statistics and Data Science Center. A better way to turn 2D designs into 3D models for rapid prototyping 23d ago “GIFT” is a new system that teaches vision-language generative AI models to produce accurate, computer-aided design (CAD) programs that can be used to simulate and test 3D objects. The method is more accurate than competing techniques, usin... 3 Questions: Neural transparency and the future of AI design 23d ago MIT Assistant Professor Pat Pataranutaporn describes a new interface that lets everyday users glimpse inside an AI's neural network before their chatbot ever says a word. Can AI build a jet engine? JARVIS Challenge tests role of AI copilots in tough-tech engineering 24d ago MIT's JARVIS Challenge (Jet-engine AI Research and Validation Intensive Sprint) is a new academic competition asking MIT students to explore whether AI can compress the design-build-test cycle so engineers can build faster and better. AI agents create virtual playgrounds to help robots get crucial training data 25d ago The “SceneSmith” system developed by MIT CSAIL researchers uses AI agents to generate lifelike scenes of indoor environments like kitchens and hotels to help robots simulate everyday chores. These 3D worlds are more realistic and diverse th... New method aims to keep kids safe from illegal AI-generated content 26d ago Researchers developed an evaluation procedure that tests generative AI models for harmful capabilities without generating outputs. This could enable auditors to identify open-source models that have been adapted to produce illegal content, ... Tiny robot boats build floating structures 29d ago FloatForm, developed at MIT, is a swarm of small aquatic robots that assemble into reconfigurable structures. It could lead to floating infrastructure that builds itself into things like a temporary platform, a market, or a stage. MIT-designed educational factory embraces modern manufacturing 30d ago The FrED factory, a low-cost desktop fiber extrusion device designed and assembled by students in an educational factory at MIT, is changing how manufacturing is taught in Mexico through a collaboration with Tecnológico de Monterrey. Q&A: What is agentic AI today, and what do we want it to be? 38d ago MIT Associate Professor Phillip Isola explains what agentic AI is, how these systems are used, what applications they are best suited for, and what the future may hold for this exploding technology.
20 loaded
MR
Microsoft Research
4d ago · 10 items
Orchard: An open framework for scalable agentic AI 4d ago Orchard is an open-source framework for the research community to train and evaluate AI agents across task types. It reduces complexity while supporting strong performance from smaller models by enabling researchers to reuse the same infras... Echoverse: Deep, evolving environments for computer-use agents 8d ago Computer-use AI agents struggle with multi-step workflows like email and customer support. Echoverse trains agents in realistic environments rather than simply providing more training tasks, helping them improve as the tasks, tests, and env... EvoLib: Turning experience into evolving knowledge 8d ago LLMs do not get smarter just by remembering more. EvoLib turns experience into evolving knowledge, taking reusable skills and insights that help models learn and adapt across tasks long after deployment. Verifying Rust cryptography in SymCrypt, from standards to code 25d ago Cryptographic code supports vital protections in modern computing systems. Learn how a new method helps verify code as developers write it while preserving speed and adaptability as it gets implemented and evolves: Aurora 1.5: Extending open foundation models for weather and Earth-system applications 29d ago Aurora 1.5 adds 22 more variables, hourly temporal resolution, and probabilistic ensemble forecasting to the Aurora foundation model, making it more useful for real-world weather, climate, and energy applications. Flint: A visualization language for the AI era 30d ago Short chart specifications are easy to write, but often produce uninspiring results. Flint is an open-source visualization language that offers a middle path, letting AI agents create expressive charts from compact, human-editable specifica... SkillOpt: Agent skills as trainable parameters 38d ago AI agents often fail because their instructions, or skills, are manually modified with no guarantee of improvement. Learn how SkillOpt turns skill editing into a training process, making agent behavior more reliable without changing model w... Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity 39d ago AI agents can't remember past conversations. They must constantly reload or retrieve context, which grows less efficient as tasks get longer and more complex. Memora solves this with a scalable memory system separating what’s stored from ho... Understanding the brain with AI-driven explanations and experiments 43d ago Researchers introduce generative causal testing, which translates black box models into clear hypotheses and verifies them in the scanner, revealing what specific brain regions respond to in language. Talos: Scaling rare disease diagnosis with automated, iterative genomic reanalysis 44d ago Talos was built to help resolve a major bottleneck in genomic medicine: human review time. The open-source system recovered 90% of in-scope diagnoses while surfacing just 1.3 candidate variants per patient for expert review.
IA
Interconnects AI
4d ago · 20 items
20 loaded
IA
Import AI
4d ago · 20 items
Import AI 467: Self-sustaining AI viruses; pacing AI progress; confusion about AI and creativity 4d ago Import AI 466: The bitter lesson for robotics, AIs complete week-long programming tasks; and OpenAI's accidental AI hacker 11d ago Import AI 465: Open vs closed gaps; Kimi K3; Demis' big policy plan 18d ago Import AI 464: Fable writes GPU kernels; AI automation; and analog computation 32d ago Import AI 463: Self-improving robots; a 10k Chinese GPU cluster; and an elegiac essay for the human era 39d ago Import AI 462: Superpersuasion; self-sustaining AI; paths to ASI 46d ago Import AI 461: "Alignment is not on track"; FrontierCode; and synthetic research interns 53d ago Import AI 460: Reward hacking society, RSI data from Anthropic; and RL-based quadcopter racing 60d ago Import AI 459: AI oversight is difficult; scaling laws for protein folding models; and pricing the extinction risk of AI systems 67d ago Import AI 458: Reckoning with the future; and a singularity story 73d ago
20 loaded
LD
Linear Digressions
5d ago · 20 items
Reasoning Models: When LLMs Went Beyond Fancy Autocomplete 5d ago Reasoning models don't just answer your question — they *think out loud* first. In this episode we dig into the class of AI models that generate intermediate chains of thought before arriving at a final answer, exploring how the internal re... Distillation, or, How to Steal a Model 12d ago This week we’re covering model distillation: the technique of using a large "teacher" model's outputs to train a smaller, cheaper "student" model that mimics it. They cover the two big reasons labs do this — making lighter, faster, more foc... Invisible LLM Failures and AI Fluency with Chris Potts (Stanford) 18d ago What happens when a Stanford linguistics professor turns his attention to AI chatbots — and the surprisingly invisible ways humans misunderstand them? Chris Potts joins the show to unpack the hidden failure modes in how we interact with AI,... Interviewing the Linear Digressions Agents (The Agents Season, Episode 11) 40d ago After a five-year hiatus, the podcast that burned out partly over the tedium of writing episode descriptions is back — and using AI agents to handle exactly that task. The season-11 finale turns the lens on the podcast itself, putting the A... Agent Economics (The Agents Season, Episode 10) 47d ago What if building more highways made your commute *slower*? That's the paradox at the heart of AI agent economics: even as per-token inference costs have plummeted dramatically over the past two years, total LLM spending keeps climbing. Draw... Agent Trust, Oversight and Control (The Agents Season, Episode 9) 54d ago Capabilities get all the attention when it comes to AI agents — but what happens when a highly capable agent makes a bad decision in the real world? Trust, oversight, and control are the unglamorous but critically important flip side of the... Many Agents, Many Problems (The Agents Season, Episode 8) 61d ago Whether you work best solo or thrive in a team, you know collaboration is complicated — and it turns out AI agents face the same tensions. This episode dives into multi-agent systems, exploring how networks of AI agents can overcome the ind... How Do You Evaluate An AI Agent? (The Agents Season, Episode 7) 68d ago Knowing when an AI agent has failed sounds straightforward — until it isn't. Agents have a frustrating habit of finishing confidently while quietly doing the wrong thing, or looping endlessly without ever crashing in an obvious way. This ep... AI Agent Failure Modes (The Agents Season, Episode 6) 74d ago Despite what the marketing hype might suggest, AI agents are far from infallible — and if you've ever actually used one, you already know this. Today's episode dives deep into the many, varied, and sometimes surprising ways AI agents can fa... Agentic Planning (The Agents Season, Episode 5) 82d ago When tackling a complex, multi-step task, even the smartest AI agent can fail without a solid game plan. This episode dives into the research around agentic planning — how agents move beyond simply reacting to what's in front of them and in...
20 loaded
ML
Machine Learning Street Talk
7d ago · 15 items
15 loaded
TL
The latest research from Google
8d ago · 20 items
20 loaded
LF
Lex Fridman Podcast
10d ago · 20 items
#499 – Gary Gallagher: American Civil War, Slavery, Lincoln, Grant & Lee 10d ago Gary Gallagher is a historian of the American Civil War. Thank you for listening ❤ Check out our sponsors: https://lexfridman.com/sponsors/ep499-sc See below for timestamps, transcript, and to give feedback, submit questions, contact Lex, e... #498 – Anthony Kaldellis: Roman Empire, Byzantine Empire, Rise & Fall of Empires 38d ago Anthony Kaldellis is a historian of the Roman Empire and author of “The New Roman Empire”, a comprehensive history of the Byzantine Empire (Eastern Roman Empire). Thank you for listening ❤ Check out our sponsors: https://lexfridman.com/spon... #497 – Biggest Mysteries in Physics: Antimatter, Dark Energy & ToE – Don Lincoln 70d ago Don Lincoln is a particle physicist at Fermilab who has spent decades working at the frontiers of high energy physics. Thank you for listening ❤ Check out our sponsors: https://lexfridman.com/sponsors/ep497-sc See below for timestamps, and ... #496 – FFmpeg: The Incredible Technology Behind Video on the Internet 93d ago Jean-Baptiste Kempf is lead developer of VLC and president of VideoLAN. Kieran Kunhya is a longtime FFmpeg contributor, codec engineer, and the person behind the now-infamous FFmpeg account on X. Thank you for listening ❤ Check out our spon... #495 – Vikings, Ragnar, Berserkers, Valhalla & the Warriors of the Viking Age 120d ago Lars Brownworth is a historian, teacher, podcaster, and author specializing in Viking history, medieval Europe, and the Byzantine Empire. Thank you for listening ❤ Check out our sponsors: https://lexfridman.com/sponsors/ep495-sc See below f... #494 – Jensen Huang: NVIDIA – The $4 Trillion Company & the AI Revolution 137d ago Jensen Huang is the co-founder and CEO of NVIDIA, the world’s most valuable company and the engine powering the AI computing revolution. Thank you for listening ❤ Check out our sponsors: https://lexfridman.com/sponsors/ep494-sc See below fo... #493 – Jeff Kaplan: World of Warcraft, Overwatch, Blizzard, and Future of Gaming 149d ago Jeff Kaplan is a legendary Blizzard game designer of World of Warcraft and Overwatch, now preparing to launch a new game, The Legend of California, from his new studio Kintsugiyama – available to wishlist on Steam today, with alpha later in... #492 – Rick Beato: Greatest Guitarists of All Time, History & Future of Music 160d ago Rick Beato is a music educator, interviewer, producer, songwriter, and a true multi-instrument musician, playing guitar, bass, cello & piano. His incredible YouTube channel celebrates great musicians & musical ideas, and helps millions of p... #491 – OpenClaw: The Viral AI Agent that Broke the Internet – Peter Steinberger 177d ago Peter Steinberger is the creator of OpenClaw, an open-source AI agent framework that’s the fastest-growing project in GitHub history. Thank you for listening ❤ Check out our sponsors: https://lexfridman.com/sponsors/ep491-sc See below for t... #490 – State of AI in 2026: LLMs, Coding, Scaling Laws, China, Agents, GPUs, AGI 188d ago Nathan Lambert and Sebastian Raschka are machine learning researchers, engineers, and educators. Nathan is the post-training lead at the Allen Institute for AI (Ai2) and the author of The RLHF Book. Sebastian Raschka is the author of Build ...
20 loaded
LF
Lex Fridman
10d ago · 15 items
Gary Gallagher: American Civil War, Slavery, Lincoln, Grant & Lee | Lex Fridman Podcast #499 10d ago The Rise and Fall of the Roman Empire and the Byzantine Empire | Lex Fridman Podcast #498 38d ago Biggest Mysteries in Physics: Antimatter, Dark Energy & ToE - Don Lincoln | Lex Fridman Podcast #497 70d ago FFmpeg: The Incredible Technology Behind Video on the Internet | Lex Fridman Podcast #496 93d ago Vikings, Ragnar, Berserkers, Valhalla & the Warriors of the Viking Age | Lex Fridman Podcast #495 120d ago Jensen Huang: NVIDIA - The $4 Trillion Company & the AI Revolution | Lex Fridman Podcast #494 137d ago Jeff Kaplan: World of Warcraft, Overwatch, Blizzard, and Future of Gaming | Lex Fridman Podcast #493 149d ago Lex trains w/ Khabib Nurmagomedov | Exclusive Footage at UFC PI 157d ago Rick Beato: Greatest Guitarists of All Time, History & Future of Music | Lex Fridman Podcast #492 160d ago Khabib vs Lex: Training with Khabib | FULL EXCLUSIVE FOOTAGE 163d ago
15 loaded
3B
3Blue1Brown
14d ago · 15 items
15 loaded
OU
One Useful Thing
15d ago · 20 items
20 loaded
WL
Welch Labs
16d ago · 19 items
19 loaded
SE
sentdex
16d ago · 15 items
15 loaded
SR
Salmon Run
19d ago · 20 items
Book Review: Domain Specific Small Language Models 19d ago Artificial Intelligence (AI) powered applications are changing the way we consume and use information. Retrieval Augmented Generation (RAG) ... Book Review: Software Engineering for Data Scientists 187d ago As a Software Engineer (backend Web Development then Search) turned Data Scientist, I was particularly interested in what the book Software ... Book Review: Transformers In Action 209d ago The Attention Is All You Need paper proposed the Transformer Architecrture as an improvement to the dominant encoder-decoder models of the ... Trip Report: PyData Global 2025 224d ago I attended PyData Global 2025 earlier this month. I had hoped to write this up earlier, but I've been busy, so only now getting the time Ch... Book Review: Time Series Forecasting using Foundation Models 299d ago As someone who primarily works in NLP and Search in the Health Domain, I don't have much use for Time Series. However, while exploring the F... Book Review: Statistics every Programmer Needs 321d ago I recently read Statistics every Programmer Needs by Gary Sutton. I am probably a good target audience for the book since I used to be a so... Book Review: Hands-On Artificial Intelligence for IoT 405d ago For those in similar professional circles as I am in, i.e. looking forward into the Generative AI space, yet with one foot pragmatically and... Book Review: Essential Graph RAG 418d ago Coming from a background of Knowledge Graph (KG) backed Medical Search, I don't need to be convinced about the importance of manually curate... Packaging ML Pipelines from Experiment to Deployment 584d ago As an ML Engineer, we are generally tasked with solving some business problem with technology. Typically it involves leveraging data assets ... Trip Report - PyData Global 2024 606d ago I attended PyData Global 2024 last week. Its a virtual conference, so I was able to attend it from the comfort of my home, although presenta...
20 loaded
AO
Ahead of AI
20d ago · 20 items
20 loaded
ON
OpenAI News
29d ago · 20 items
20 loaded
How Open Models Are Driving AI Research 32d ago NVIDIA open models from Nemotron, Cosmos and BioNeMo are fueling the field's biggest research questions at ICML 2026. NVIDIA Research Unlocks Advanced Grasping, Smarter Autonomous Driving and Agent Training at Scale 65d ago New NVIDIA Research breakthroughs show how training at scale — across gripper types, driving scenarios and virtual worlds — creates AI that generalizes to diverse applications. NVIDIA Enables the Next Era Of Physical AI Research With Agent Skills For Autonomous Vehicles, Robotics And Vision AI 65d ago New physical AI agent skills, powered by NVIDIA Cosmos 3, help researchers accelerate data generation, simulation, policy training and evaluation for autonomous system development. NVIDIA Research Advances Robotics From Simulation to the Real World 71d ago Featured at the International Conference on Robotics and Automation, eight new NVIDIA Research papers show how robots trained in simulation are moving into the real world. NVIDIA Launches Earth-2 Family of Open Models — the World’s First Fully Open, Accelerated Set of Models and Tools for AI Weather 193d ago NVIDIA Earth-2 makes weather AI accessible worldwide at every stage — from processing initial observation data to generating 15-day global forecasts or local storm forecasts. At NeurIPS, NVIDIA Advances Open Model Development for Digital and Physical AI 249d ago NVIDIA releases new AI tools for speech, safety and autonomous driving — including NVIDIA DRIVE Alpamayo-R1, the world’s first open industry-scale reasoning vision language action model for mobility — and a new independent benchmark recogni... How Do You Teach an AI Model to Reason? With Humans 345d ago NVIDIA’s data factory team creates the foundation for AI models like Cosmos Reason, which today topped the physical reasoning leaderboard on Hugging Face. NVIDIA Research Shapes Physical AI 361d ago AI and graphics research breakthroughs in neural rendering, 3D generation and world simulation power robotics, autonomous vehicles and content creation. NVIDIA Research Showcases the Future of Robotics at RSS 413d ago At this year’s Robotics: Science and Systems conference, NVIDIA Research is presenting work that advances robot learning across simulation, real-world transfer and decision-making. NVIDIA Scores Consecutive Win for End-to-End Autonomous Driving Grand Challenge at CVPR 422d ago NVIDIA was today named an Autonomous Grand Challenge winner at the Computer Vision and Pattern Recognition (CVPR) conference, held this week in Nashville, Tennessee. The announcement was made at the Embodied Intelligence for Autonomous Syst...
18 loaded
LL
Lil'Log
35d ago · 20 items
Harness Engineering for Self-Improvement 35d ago The concept of recursive self-improvement (RSI) dates back to I. J. Good (1965), where he defined an “ultraintelligent machine” as a system that can surpass humans in all intellectual activities and design better machines to improve itself.... Scaling Laws, Carefully 45d ago Scaling laws are one of the most critical empirical findings in deep learning. The observation is simple in form: the training loss $L$ decreases predictably as we scale up model size $N$, dataset size $D$, and compute $C$, following a powe... Why We Think 464d ago Special thanks to John Schulman for a lot of super valuable feedback and direct edits on this post. Test time compute (Graves et al. 2016, Ling, et al. 2017, Cobbe et al. 2021) and Chain-of-thought (CoT) (Wei et al. 2022, Nye et al. 2021), ... Reward Hacking in Reinforcement Learning 618d ago Reward hacking occurs when a reinforcement learning (RL) agent exploits flaws or ambiguities in the reward function to achieve high rewards, without genuinely learning or completing the intended task. Reward hacking exists because RL enviro... Extrinsic Hallucinations in LLMs 762d ago Hallucination in large language models usually refers to the model generating unfaithful, fabricated, inconsistent, or nonsensical content. As a term, hallucination has been somewhat generalized to cases when the model makes mistakes. Here,... Diffusion Models for Video Generation 848d ago Diffusion models have demonstrated strong results on image synthesis in past years. Now the research community has started working on a harder task—using it for video generation. The task itself is a superset of the image case, since an ima... Thinking about High-Quality Human Data 915d ago [Special thank you to Ian Kivlichan for many useful pointers (E.g. the 100+ year old Nature paper “Vox populi”) and nice feedback. 🙏 ] High-quality data is the fuel for modern data deep learning model training. Most of the task-specific lab... Adversarial Attacks on LLMs 1018d ago The use of large language models in the real world has strongly accelerated by the launch of ChatGPT. We (including my team at OpenAI, shoutout to them) have invested a lot of effort to build default safe behavior into the model during the ... LLM Powered Autonomous Agents 1142d ago Building agents with LLM (large language model) as its core controller is a cool concept. Several proof-of-concepts demos, such as AutoGPT, GPT-Engineer and BabyAGI, serve as inspiring examples. The potentiality of LLM extends beyond genera... Prompt Engineering 1242d ago Prompt Engineering, also known as In-Context Prompting, refers to methods for how to communicate with LLM to steer its behavior for desired outcomes without updating the model weights. It is an empirical science and the effect of prompt eng...
20 loaded
EY
Eugene Yan
48d ago · 20 items
20 loaded
FO
Future of Life Institute
49d ago · 20 items
Should AIs be people too? 49d ago Statement: Anthropic warns of AI self-improvement risks, considers a pause 60d ago FLI President on the White House Executive Order 65d ago Magnificent Humanity – The Pope’s First Encyclical Concerns AI 79d ago White House working group on AI – Statement from FLI’s Anthony Aguirre 94d ago FLI’s President and CEO on Trump’s support for an AI ‘kill switch’ 113d ago FLI CEO’s statement on the attack against Sam Altman’s home 119d ago Prominent Scientists, Faith Leaders, Policymakers and Artists Call for a Prohibition on Superintelligence, as Poll Shows Americans Don’t Want It 133d ago Statement: Head of US Policy on the White House AI legislative recommendations 138d ago Governor DeSantis Directs Florida State Agencies to Partner with Future of Life Institute to Shield Families from AI Harm 151d ago
20 loaded
DS
David Stutz
54d ago · 10 items
Domain-Specific AI Should Focus on Workflows Rather Than Modeling 54d ago I spent the past few years working on AI for health. Starting with custom multimodal encoders, post-training, and sophisticated multi-agent architecturs, I now see modeling work becoming less and less important for domain-specific applicati... AI Evaluation is Becoming an Exciting Standalone Discipline 113d ago Having worked on robustness problems during my PhD, I see many of the characteristics appearing in the evaluation of LLMs and AI systems. Adversarial attacks such as jailbreaks are becoming more relevant, edge cases finally become relevant,... RAISE 2025 Panel Statement on Aligning AI to Clinical Values 307d ago Recently, I attended the Responsible AI for Social and Ethical Healthcare 2025 “2.0” Symposium organized by, among others, Harvard Medical School. The symposium featured various panels on topics surrounding generative AI, in particular mult... Some Lessons on Reviews and Rebuttals 551d ago Writing and responding to reviews is the bread and butter of any academic and especially in AI research, PhD students are confronted with both rather early compared to other displicines. Unfortunately, I found that drafting reviews and rebu... Thoughts on Watermarking AI-Generated Content 569d ago Watermarking AI-generated content has the potential to address various problems that generative AI threatens to aggravate — misinformation, impersonation, copyright infringement, web pollution, etc. However, it is also controversial with ma... Thoughts and Lessons for Planning Rater Studies in AI 578d ago With the goal of deploying generative AI systems, rater studies are becoming increasingly common and important. This means more and more researchers and engineers face the challenge of actually planning and conducting rater studies for AI s... Open-Sourcing Relabeled MedQA and Dermatology DDx Datasets 633d ago Dealing with rater disagreement is becoming more important in AI, especially for LLMs and in specialized domains such as health. In the past year, I helped open source two datasets allowing to study rater disagreement in the health domain: ... Thinking About Research Ideas vs. Technology 635d ago In this article, I want to share some thoughts on the difference between research ideas and technology, particularly in machine learning. This distinction is have been contemplating since starting my PhD. After joining Google DeepMind and b... The Importance of Effectively Experimenting in an AI PhD 760d ago Engineering and running experiments are a key component of most PhDs in AI. While there are plenty of more theoretical topics that are often limited to smaller scale experimentation, the trend has definitely been to scale up models, dataset... FAQ for our Monte Carlo Conformal Prediction 817d ago Over the past months, I have given several talks about Monte Carlo conformal prediction and the problem of calibrating with uncertain ground truth, for example, stemming from annotator disagreement. Each time, the audience had great questio...
LM
Learn Machine Learning
57d ago · 2235 items
We are open-sourcing an vision-language interaction model and system 57d ago We are open-sourcing an vision-language interaction model and system 57d ago We are open-sourcing an vision-language interaction model and system 57d ago We are open-sourcing an vision-language interaction model and system 57d ago Everyone's been talking about Thinking Machines' "interaction model" as a concept. We went and built one — 8B, vision-driven, decides on its own when to speak, and when to delegate to agent — and we're open-sourcing all of it 57d ago Math Focused Learning Resources 57d ago I built a FIFA World Cup 2026 Predictive Oracle & Bracket Simulator live on the web! 57d ago How is CampusX One Membership courses ?? worth it? 57d ago GitHub Autopilot — Open Source GitHub App for Repository Automation 57d ago Day 20 of Reviewing 1 free AI, ML, or data certification every day, so you don’t have to waste time with bad courses. 57d ago
2235 loaded
AI
Artificial Intelligence (AI)
57d ago · 1197 items
[ Removed by Reddit ] 57d ago Visa and OpenAI Let AI Agents Shop on Your Behalf Using Visa's Global Network 57d ago The productivity gap between "AI user" and "AI agent user" is bigger than I expected 57d ago If AI could comfort you perfectly, would you still want parts of yourself left unread? 57d ago How To Get Web Design Clients 57d ago Which AI agent are you? 57d ago Do you think AI is becoming normal faster than people expected? 57d ago By 2050, we may see AI assistants in every home, personalized learning for every student, advanced medical treatments, smart cities, and even human-AI collaboration on a massive scale. 57d ago The gap between decision and exécution 57d ago OpenAI Filed for IPO at $852B as Anthropic Beats It to Market and Price Cuts Loom 57d ago
1197 loaded
DL
Deep Learning
57d ago · 610 items
SenseNova U1 training code and dataset are open-sourced. How is it different from other text-to-image models? 57d ago JudgeOS V5.7 / EBH — The Governance Firewall Above AI, Robots, Agents, and Autonomous Workflows 57d ago “GenalShift (mi función de activación) ha superado a ReLU en CIFAR-10 entrenando una ResNet18 desde cero: 92.33% vs 92.07% (+0.26%). Código abierto en GitHub. #IAsoberana #DeepLearning” 57d ago I spent a year applying information geometry to LLM behavioral monitoring. Here’s what the math shows about multi-turn attacks. 57d ago Nobody sent a memo. Nobody made an announcement. But somewhere in the last two years, marketing quietly changed forever. 57d ago Need help with implementation of transformer-decoder model 57d ago Plot twist: your future killer already has a USB port 57d ago Running Gemma 4 QAT 12B on an 8GB GPU at 16k context — measured the KV-cache tradeoffs 58d ago Request for critique: deterministic governance boundary for AI agent actions before execution 58d ago Analysis of the results of the "Transforming autoencoders" architecture mentioned by Hilton, for my dissertation. 58d ago
610 loaded
RL
Reinforcement Learning
57d ago · 254 items
highway-v0 env is too slow 57d ago Fair Reinforcement Learning 57d ago Korrel: turn one agent eval into a verifiers or OpenEnv RL environment, with a fidelity proof against tau2-bench 58d ago Do you ever get to the point of mental breakdown? 58d ago Roast my resume 58d ago Optimizing an RL Training Pipeline: Memory, Sampling, and Copy Elimination 58d ago Resoning LLMs make RL agent learn Faster 58d ago I Built a Reinforcement Learning AI That Runs on an Arduino Mega 58d ago Previous Claude models struggled to play Pokémon Fire even with harnesses that gave them additional helpful tools, but Fable 5 beat FireRed with a minimal, vision-only harness. 59d ago Testing the stability of my new walking gait (x0.25) 59d ago
254 loaded
DS
Data Science
57d ago · 141 items
Is this AgenticAI Ragebait? 57d ago How to stop shipping low-quality RL environments, with examples 58d ago Is my tech stack becoming a liability for future job prospects? 58d ago AI Overuse Follow-up 59d ago How do you measure to performance / accuracy of a recommender system? 59d ago How do you put a price on a healthy work environment and a good manager? 59d ago How Earnings Impact My Momentum Strategy - A Backtest Across Two XGBoost Models 59d ago What Data Structures and Algorithms topics actually come up in technical interviews? 59d ago Weekly Entering & Transitioning - Thread 08 Jun, 2026 - 15 Jun, 2026 61d ago Open and closed models are on different exponentials 61d ago
141 loaded
ONLY FOR A LIMITED TIME 58d ago Designing a Universal AI Safety Framework 58d ago Analog Neuromorphic letter recognition circuit 58d ago How one engineer at Spotify solved the recommendations of music by building an open source library ANNOY 58d ago Built native iOS/macOS ONNX model analyzer (inspired by Netron) (Looking for feedback) 58d ago Optimizing an RL Training Pipeline: Memory, Sampling, and Copy Elimination 58d ago Personalization Yo-Yo: A Ruler-Based Mechanism for Non-Sticky Long-Term Personalization 60d ago Object detection Using Detection Transformer (Detr) for Bone fraction dataset 61d ago I built an MNIST classifier from scratch in pure Python (no NumPy) to actually understand backprop 63d ago dataset and architecture 64d ago
59 loaded
PA
Partners and Integrations
80d ago · 10 items
IN
inFERENCe
163d ago · 15 items
The Future of Software 163d ago The world of software is undergoing a shift not seen since the advent of compilers in the 1970s. Compilers were the original vibe coding: they automatically generate complex machine code that human programmers had to manually write before. ... Deep Learning is Powerful Because It Makes Hard Things Easy - Reflections 10 Years On 188d ago Ten years ago this week, I wrote a post called "Deep Learning is Easy - Learn Something Harder". The post blew up, top spot on HackerNews. Needless to say, it didn't age well. Discrete Diffusion: Continuous-Time Markov Chains 442d ago A tutorial explaining some intuitions behind continuous time Markov chains for machine learners interested in discrete diffusion models. We may finally crack Maths. But should we? 1156d ago Automating mathematical theorem proving has been a long standing goal of artificial intelligence and indeed computer science. It's one of the areas I became very interested in recently. This is because I feel we may have the ingredients nee... Mortal Komputation: On Hinton's argument for superhuman AI. 1165d ago Last week in Cambridge was Hinton bonanza. He visited the university town where he was once an undergraduate in experimental psychology, and gave a series of back-to-back talks, Q&A sessions, interviews, dinners, etc. He was stopped on the ... Autoregressive Models, OOD Prompts and the Interpolation Regime 1226d ago A few years ago I was very much into maximum likelihood-based generative modeling and autoregressive models (see this, this or this). More recently, my focus shifted to characterising inductive biases of gradient-based optimization focussin... We May be Surprised Again: Why I take LLMs seriously. 1234d ago "Deep Learning is Easy, Learn something Harder" - I proclaimed in one of my early and provocative blog posts from 2016. While some observations were fair, that post is now evidence that I clearly underestimated the impact simple techniques ... Implicit Bayesian Inference in Large Language Models 1618d ago This intriguing paper kept me thinking long enough for me to I decide it's time to resurrect my blogging (I started writing this during ICLR review period, and realised it might be a good idea to wait until that's concluded) * Sang Michael ... Eastern European Guide to Writing Reference Letters 1621d ago Excruciating. One phrase I often use to describe what it's like to read reference letters for Eastern European applicants to PhD and Master's programs in Cambridge. Even objectively outstanding students often receive dull, short, factual, a... Causal inference 4: Causal Diagrams, Markov Factorization, Structural Equation Models 1884d ago This post is written with my PhD student and now guest author Patrik Reizinger [https://twitter.com/rpatrik96] and is part 4 of a series of posts on causal inference: * Part 1: Intro to causal inference and do-calculus [https://www.inferenc...
15 loaded
DB
Damian Bogunowicz - dtransposed
169d ago · 10 items
TG
The Gradient
170d ago · 15 items
15 loaded
AK
Andrej Karpathy blog
176d ago · 10 items
VI
VITALab
214d ago · 10 items
Towards Brain MRI Foundation Models for the Clinic: Findings from the FOMO25 Challenge 214d ago 1. Motivation Brain Latent Progression Individual-based spatiotemporal disease progression on 3D Brain MRIs via latent diffusion 345d ago This article aims at reviewing a Alzheimer’s spatiotemporal disease progression predictive model called Brain Latent Progression (BrLP). All in all, this is ... A Survey of popular LLM Evaluation Metrics 353d ago Large Language Models (LLMs) are increasingly applied to critical domains such as medical report generation, where accuracy and trust are essential. Evaluati... Open-Source Large Language Models in Radiology: A Review and Tutorial for Practical Research and Clinical Deployment 362d ago Open-Source Large Language Models in Radiology MemSAM: Taming Segment Anything Model for Echocardiography Video Segmentation 431d ago MemSAM Simplifying Deep Temporal Difference Learning 488d ago tl;dr The authors propose PQN, a simplified deep online Q-Learning that uses very small replay buffers. Normalization and parallelized sampling from vectoriz... EchoPrime: Multi-Video View-Informed Vision-Language Model for Comprehensive Echocardiography Interpretation 502d ago Objective EchoPrime is a foundation model designed for comprehensive echocardiographic interpretation. Unlike previous models that use single views or static... DeepSeek-V3 Technical Report 543d ago DeepSeek-V3 Variational Autoencoders for Generating Synthetic Tractography-Based Bundle Templates in a Low-Data Setting 571d ago Highlights Implicit neural representations 599d ago Implicit neural networks
DA
Datumbox
468d ago · 20 items
20 loaded
JA
Jay Alammar
500d ago · 10 items
Moving To Substack 500d ago I’m freezing this blog and starting to post on my Substack instead. The authoring experience is much more convenient for me there. Please follow me there, and check out The Illustrated DeepSeek R-1 if you haven’t yet. And check out our How ... Generative AI and AI Product Moats 1187d ago Here are eight observations I’ve shared recently on the Cohere blog and videos that go over them.: Article: What’s the big deal with Generative AI? Is it the future or the present? Article: AI is Eating The World Remaking Old Computer Graphics With AI Image Generation 1315d ago Can AI Image generation tools make re-imagined, higher-resolution versions of old video game graphics? Over the last few days, I used AI image generation to reproduce one of my childhood nightmares. I wrestled with Stable Diffusion, Dall-E ... The Illustrated Stable Diffusion 1404d ago Translations: Chinese, Vietnamese. (V2 Nov 2022: Updated images for more precise description of forward diffusion. A few more images in this version) AI image generation is the most recent AI capability blowing people’s minds (mine included... Applying massive language models in the real world with Cohere 1615d ago A little less than a year ago, I joined the awesome Cohere team. The company trains massive language models (both GPT-like and BERT-like) and offers them as an API (which also supports finetuning). Its founders include Google Brain alums in... The Illustrated Retrieval Transformer 1678d ago Discussion: Discussion Thread for comments, corrections, or any feedback. Translations: Korean, Russian Summary: The latest batch of language models can be much smaller yet achieve GPT-3 like performance by being able to query a database or... Explainable AI Cheat Sheet 1922d ago Introducing the Explainable AI Cheat Sheet, your high-level guide to the set of tools and methods that helps humans understand AI/ML models and their predictions. I introduce the cheat sheet in this brief video: Finding the Words to Say: Hidden State Visualizations for Language Models 2027d ago By visualizing the hidden state between a model's layers, we can get some clues as to the model's Interfaces for Explaining Transformer Language Models 2060d ago Interfaces for exploring transformer language models by looking at input saliency and neuron activation. Explorable #1: Input saliency of a list of countries generated by a language model Tap or hover over the output tokens: Explorable #2: ... How GPT3 Works - Visualizations and Animations 2203d ago Discussions: Hacker News (397 points, 97 comments), Reddit r/MachineLearning (247 points, 27 comments) Translations: German, Korean, Chinese (Simplified), Russian, Turkish The tech world is abuzz with GPT3 hype. Massive language models (lik...
AK
Andrej Karpathy
526d ago · 15 items
15 loaded
CH
Chip Huyen
569d ago · 10 items

No matching sources found.