AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,631 results
30 Jun 2026

OmniCoT: A Benchmark for Global and Multi-Step Panoramic Reasoning

Model ReleasesDGX agent

arXiv:2606.30378v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated promising spatial reasoning capabilities, while these abilities remain underexplored in the e

Online Data Selection for Instruction Tuning via Gaussian Processes

Local AiDGX agent

arXiv:2606.30077v1 Announce Type: cross Abstract: With Large Language Model (LLM) pre-training and fine-tuning shifting its focus from data volume to data quality, quality data selection has emerged a

Our first summit dedicated to world models, coming to SF this September.

IndustryDGX agent

Our first summit dedicated to world models, coming to SF this September. Announcing our first summit dedicated to world models, coming to SF this September. Very excited to have some incredible resear

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Persona-Trained Monte Carlo: Estimating Market-Outcome Distributions via Swarms of Persona-Conditioned Neural Policy Bots in a Limit Order Book

SafetyDGX agent

arXiv:2606.29556v1 Announce Type: new Abstract: We propose Persona-Trained Monte Carlo (PTMC), a method for estimating distributions of market-outcome statistics by repeatedly simulating limit-order-b

Reinforcement Learning for Software Vulnerability Analysis: A Systematic Review with Emphasis on C/C++ Source Code and Static Analysis

AgentsDGX agent

arXiv:2606.28403v1 Announce Type: cross Abstract: Vulnerability detection in C/C++ software remains a major security challenge due to code complexity, manual memory management, and the limitations of

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation

Model ReleasesDGX agent

arXiv:2606.28998v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment trains an LLM using preference data to produce outputs that better meet established quality standards. While LLM

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation

SafetyDGX agent

arXiv:2602.09305v2 Announce Type: replace Abstract: Large Language Models (LLMs) demonstrate transformative potential, yet their reasoning remains inconsistent and unreliable. Reinforcement learning (

SA-Homo: Scale Adaptive Homography Estimation for Scale Variation Scenarios

Model ReleasesDGX agent

arXiv:2606.30408v1 Announce Type: new Abstract: Homography estimation, as one of the fundamental problems in computer vision, remains challenged by scale variation scenarios where image pairs potentia

SADL: What to Ignore? A Benchmark for Subject-Aware Distractor Localization

Model ReleasesDGX agent

arXiv:2606.30393v1 Announce Type: new Abstract: Photographs frequently contain visual distractors besides foregrounds and backgrounds of the intended subject, competing for attention and weakening com

ScarfBench: Benchmarking AI Agents for Enterprise Java Framework Migration

ApplicationsDGX agent

ScarfBench is a benchmarking framework designed to evaluate AI agents' capabilities in migrating enterprise Java applications to modern frameworks. The benchmark likely assesses how well AI systems ca

Solver-Verified Formulation Generation and Selection for Multi-Warehouse Inventory Allocation Using Large Language Models

ApplicationsDGX agent

arXiv:2606.29366v1 Announce Type: cross Abstract: Balance-oriented multi-warehouse inventory allocation is a recurring decision problem in large-scale e-commerce supply chains, in which a fixed replen

StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley

Model ReleasesDGX agent

arXiv:2507.07445v3 Announce Type: replace Abstract: Autonomous agents navigating human society must master both production activities and social interactions, yet existing benchmarks rarely evaluate t

SwarmX: Agentic Scheduling for Low-Latency Agentic Systems

HardwareDGX agent

arXiv:2606.21401v2 Announce Type: replace-cross Abstract: Agentic AI applications compose multiple model calls and tool executions, creating new scheduling challenges for GPU-CPU clusters. Their infer

SWE-fficiency: Can Language Models Optimize Real-World Repositories on Real Workloads?

Model ReleasesDGX agent

arXiv:2511.06090v3 Announce Type: replace-cross Abstract: Optimizing the performance of large-scale software repositories demands expertise in code reasoning and software engineering (SWE) to reduce r

The entire business of the AI Labs & the potential speed of transformation stems from this fact.

ApplicationsDGX agent

Ethan Mollick discusses a fundamental factor driving AI Labs' business model and the rapid pace of AI transformation. The post likely explores how a core principle or technological capability underpin

The Heterogeneous Safety Impacts of Benign Multilingual Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.28843v1 Announce Type: cross Abstract: Fine-tuning a large language model is a ubiquitous method for enhancing its capability on a specific downstream task. However, prior work has shown th

The Interference Gap: Comparing Retrieval Bounds in Human Memory and RAG Systems

Model ReleasesDGX agent

arXiv:2606.28327v1 Announce Type: cross Abstract: How do retrieval bounds compare between human episodic memory and Retrieval-Augmented Generation (RAG) systems under semantic interference? We present

The US going 100% EV by 2040 would save more than 100k lives, study says

IndustryDGX agent

A study from the American Lung Association found that if the U.S. transitions to a 100% zero-emission passenger vehicle fleet by 2050, it would result in 89,300 fewer premature deaths and $978 billion

Thrilled to announce the Wearable AI Workshop at ECCV 2026 🎉 that we're organizing with an awesome group of folks across Meta Reality Labs,…

Model ReleasesDGX agent

Thrilled to announce the Wearable AI Workshop at ECCV 2026 🎉 that we're organizing with an awesome group of folks across Meta Reality Labs, AMI Labs, HKUST, Georgia Tech, UCF, and U. of Edinburgh. If

Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2606.28425v1 Announce Type: cross Abstract: Increasingly autonomous agentic AI systems pose novel multi-agent risks, such as secret collusion via covert communication channels. The natural defen

Towards Continual Motion-Language Agents: LoRA Variants for Incremental Motion Understanding and Generation

Model ReleasesDGX agent

arXiv:2606.30266v1 Announce Type: cross Abstract: Motion-language agents must possess the bidirectional capability to both understand human movement (motion-to-text, M2T) and generate it from natural

Towards Physical Intuitions for Alignment Dynamics: A Case Study With Randomness Crystallization

SafetyDGX agent

arXiv:2606.29933v1 Announce Type: new Abstract: The alignment of language models is typically studied through the lens of capability benchmarks, but the dynamics of how models change during post-train

Tutorial on using Gemini live to build a voice agent Uses deepagents as a tool: offload complex work to this subagent, use Gemini live for t…

Model ReleasesDGX agent

Tutorial on using Gemini live to build a voice agent Uses deepagents as a tool: offload complex work to this subagent, use Gemini live for the naturalness/latency Building voice agents can come with t

When Medical Safety Alignment Fails: A Benchmark for Evaluating LLMs on High-Risk Medical Queries

Model ReleasesDGX agent

arXiv:2606.28332v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for medical and health-related questions, yet their safety in high-risk medical scenarios remains p

When Stopping Fails: Rethinking Minimal Risk Conditions through Human-Interactive Autonomous Driving for Safe Transportation Systems

SafetyDGX agent

arXiv:2606.29115v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) are increasingly deployed in urban environments, yet their safety frameworks remain primarily designed around collision avoi

29 Jun 2026

A Comparison of Fusion Techniques for Multi-Modal Human Activity Recognition on the HARMES Dataset

Model ReleasesDGX agent

arXiv:2606.27886v1 Announce Type: new Abstract: Recent advances in Human Activity Recognition (HAR) from wearable sensors have shown that multi-modal deep learning models consistently outperform their

Agentic AI-Powered Re-Identification: An Emerging, Scalable Threat to Mobility Microdata Privacy

AgentsDGX agent

arXiv:2606.27936v1 Announce Type: cross Abstract: The widespread collection of fine-grained location data by commercial data brokers creates a re-identification risk that is not widely recognised by t

Agentic Hardware Design as Repository-Level Code Evolution

Model ReleasesDGX agent

arXiv:2606.28279v1 Announce Type: cross Abstract: We present HORIZON, a self-evolving agent framework that treats hardware design as repository-level code evolution. A Markdown harness is compiled int

AI leaderboard provider Arena says it hit $100M in annualized run-rate revenue eight months after launching AI Evaluations, which offers performance analytics (Marina Temkin/TechCrunch)

IndustryDGX agent

Marina Temkin / TechCrunch: AI leaderboard provider Arena says it hit $100M in annualized run-rate revenue eight months after launching AI Evaluations, which offers performance analytics — Just eight

Benchmarking Multi-Modal Graph-based Social Media Popularity Prediction

Model ReleasesDGX agent

arXiv:2606.27539v1 Announce Type: cross Abstract: Social media popularity prediction aims to forecast the future reach or influence of online content from early-stage observations. Accurate prediction

Complex-Valued 2D Gaussian Representation for Computer-Generated Holography

Model ReleasesDGX agent

arXiv:2511.15022v2 Announce Type: replace Abstract: Complex-valued Gaussian primitives have recently been explored for representing holographic radiance fields in 3D novel view synthesis. In this work

Continual Learning for Sequential Personalization of Small Language Models: A Stability Monitoring Analysis

Local AiDGX agent

arXiv:2606.27634v1 Announce Type: new Abstract: Small Language Models (SLMs) are increasingly being considered for deployment on edge devices such as laptops, enabling private, low-latency, and locall

Data Scaling Laws in Imitation Learning for Robotic Manipulation

SafetyDGX agent

arXiv:2410.18647v4 Announce Type: replace Abstract: Data scaling has revolutionized fields like natural language processing and computer vision, providing models with remarkable generalization capabil

DeLux: Cross-Modal Local Artifact Restoration in Video Using Neuromorphic Data

TutorialsDGX agent

arXiv:2606.27576v1 Announce Type: new Abstract: Conventional RGB cameras suffer from lighting artifacts such as flare, glare, flicker, and overexposure, leading to irrecoverable information loss that

Drifting in the Future: Stabilizing Path Following Drifting on High-Latency Vehicle Systems

SafetyDGX agent

arXiv:2606.27914v1 Announce Type: new Abstract: Autonomously controlling and handling a vehicle at and beyond its stability limit is a mathematically and computationally demanding task. Prior demonstr

GRAFT: Biological Graph and Hypergraph Benchmarks for Linked Gene Expression and Phenotypic Trait Prediction in Arabidopsis thaliana

Model ReleasesDGX agent

arXiv:2606.27413v1 Announce Type: cross Abstract: Understanding which genes control which traits in an organism remains one of the central challenges in biology. Despite significant advances in data c

Import AI 463: Self-improving robots; a 10k Chinese GPU cluster; and an elegiac essay for the human era

HardwareDGX agent

This newsletter issue covers three major topics in AI development: advances in self-improving robotic systems, details about a large-scale Chinese GPU cluster with 10,000 units for AI training, and a

Learning to Throw: Agile and Accurate Cable-Suspended Payload Delivery with a Quadrotor

SafetyDGX agent

arXiv:2606.27603v1 Announce Type: new Abstract: Quadrotors offer the agility needed to rapidly transport suspended payloads during time-critical applications, including search-and-rescue and medical d

Long-Term Prediction of Local and Global Human Motion with Occlusion Recovery

Local AiDGX agent

arXiv:2606.27900v1 Announce Type: new Abstract: Human motion describes the three-dimensional full-body movement of a person. Anticipating such motion holds significant relevance across a wide range of

MobileManiBench: Simplifying Model Verification for Mobile Manipulation

Model ReleasesDGX agent

arXiv:2602.05233v2 Announce Type: replace Abstract: Vision-language-action models have advanced robotic manipulation but remain constrained by reliance on the large, teleoperation-collected datasets d

Mosaic: A Benchmark Suite for Differentiable Physics Solvers

Model ReleasesDGX agent

arXiv:2606.27895v1 Announce Type: cross Abstract: Differentiable partial differential equation (PDE) solvers underpin solver-in-the-loop ML training, gradient-based optimal control, and inverse proble

NEW paper from Google (bookmark it) It's on advancing automated scientific review. Just pay attention to the focus on agentic verification w…

AgentsDGX agent

NEW paper from Google (bookmark it) It's on advancing automated scientific review. Just pay attention to the focus on agentic verification which is something I've been writing about recently. AI is ac

Odyssey: Constructing Verifiable Local Truth-Preserving Foundation Models

Local AiDGX agent

arXiv:2606.27593v1 Announce Type: new Abstract: We introduce a categorical framework called ODYSSEY for constructing verifiable, local truth-preserving foundation models as compositions of foundries:

Position: The Term 'Machine Unlearning' Is Overused in LLMs

SafetyDGX agent

arXiv:2606.27379v1 Announce Type: cross Abstract: Large language models increasingly face demands to 'forget' training data, knowledge, or behaviors due to regulatory deletion obligations, copyright/l

Ranking Before Serving: Low-Latency LLM Serving via Pairwise Learning-to-Rank

ApplicationsDGX agent

arXiv:2510.03243v3 Announce Type: replace-cross Abstract: Efficient scheduling of large language model (LLM) inference tasks is critical for achieving low latency and high throughput, a challenge that

Room for Error: Large-Scale Simulation of Over-the-Air Acoustic Attacks

ApplicationsDGX agent

arXiv:2606.27701v1 Announce Type: cross Abstract: While voice control is rapidly becoming a ubiquitous vector of human-AI communication, the risks facing these systems remain poorly understood. This i

Seven Security Challenges That Must be Solved in Cross-domain Multi-agent LLM Systems

SafetyDGX agent

arXiv:2505.23847v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are rapidly evolving into autonomous agents that cooperate across organizational boundaries, enabling joint disas

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery

Model ReleasesDGX agent

arXiv:2505.10764v4 Announce Type: replace Abstract: Innovations in digital intelligence are transforming robotic surgery with more informed decision-making. Real-time awareness of surgical instrument

Swarm sign language: motion-based communication between drones

AgentsDGX agent

arXiv:2606.27883v1 Announce Type: new Abstract: In stealth-constrained swarm robotics, visual communication provides a critical alternative to active radio transmissions, which might be jammed. This r

The USG launching models on Hugging Face. Go @jgebbia

IndustryDGX agent

The U.S. Government is launching machine learning models on Hugging Face, a popular open-source platform for sharing AI models and datasets. This initiative, highlighted by Hugging Face co-founder Cle

Towards Automating Scientific Review with Google's Paper Assistant Tool

Model ReleasesDGX agent

arXiv:2606.28277v1 Announce Type: cross Abstract: Artificial intelligence is driving a revolution in scientific discovery, accelerating everything from hypothesis generation to mathematical theorem pr

When Does Personality Composition Matter for Multi-Agent LLM Teams?

AgentsDGX agent

arXiv:2606.27443v1 Announce Type: new Abstract: Personality prompting shapes how large language models communicate, yet whether these behavioral shifts affect objective task outcomes remains under-exp

When Multi-Robot Systems Meet Agentic AI:Towards Embodied Collective Intelligence

AgentsDGX agent

arXiv:2606.27929v1 Announce Type: new Abstract: Embodied AI is increasingly becoming agentic, shifting robots from perception--control pipelines towards closed-loop systems that can retrieve context,

ZooClaw-FashionSigLIP2: Distilled Fine-tuning for Robust Fashion Retrieval

Model ReleasesDGX agent

arXiv:2606.27708v1 Announce Type: new Abstract: Adapting a foundation vision-language encoder to a specialized retrieval task creates a fundamental tradeoff: gains on the target distribution come at t

27 Jun 2026

An interesting way to take Noam at his word in regards to always keeping a constant inference budget for any eval reporting - is that open m…

ToolsDGX agent

An interesting way to take Noam at his word in regards to always keeping a constant inference budget for any eval reporting - is that open models have a lot more dollar per token mileage than closed m

btw we crossed our 6k attendee mark a while ago. will probably call sold out when we hit 7k this weekend. do get tix now, this is the epicen…

ApplicationsDGX agent

btw we crossed our 6k attendee mark a while ago. will probably call sold out when we hit 7k this weekend. do get tix now, this is the epicenter of ai next week. if you are a student or between jobs, h

26 Jun 2026

1/ On p (doom) tl;dr a) Everyone is making up the numbers b) nobody knows anything (least of all the experts), c) don't worry about it d) th…

Model ReleasesDGX agent

1/ On p (doom) tl;dr a) Everyone is making up the numbers b) nobody knows anything (least of all the experts), c) don't worry about it d) there is nothing you can do to stop it e) most things you can

A Multi-Layer AI Framework for Information Landscape Analysis

SafetyDGX agent

arXiv:2606.26115v1 Announce Type: cross Abstract: This paper proposes a multi-layer AI framework for information landscape analysis in the context of information disorder. Rather than treating misinfo

Adversarial Robustness of AI-Generated Image Detectors in the Real World

ApplicationsDGX agent

arXiv:2410.01574v4 Announce Type: replace Abstract: The rapid advancement of Generative Artificial Intelligence (GenAI) capabilities is accompanied by a concerning rise in its misuse. In particular th

Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents

Model ReleasesDGX agent

arXiv:2606.26627v1 Announce Type: cross Abstract: Large language model agents increasingly query databases, search document collections, call external APIs, remember past interactions, and act on a us

← Previous
1…386387388389390…428
Next →