AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,629 results
21 May 2026

DriveMA: Rethinking Language Interfaces in Driving VLAs with One-Step Meta-Actions

Model ReleasesDGX agent

arXiv:2605.21273v1 Announce Type: new Abstract: Driving Vision-Language-Action Models (Driving VLAs) commonly introduce natural-language reasoning as an intermediate interface for end-to-end planning,

Findings of the Counter Turing Test: AI-Generated Image Detection

ApplicationsDGX agent

arXiv:2605.20787v1 Announce Type: new Abstract: The rapid advancements in generative AI technologies, such as Stable Diffusion, DALL-E, and Midjourney, have significantly transformed the creation of s

Findings of the Counter Turing Test: AI-Generated Text Detection

Model ReleasesDGX agent

arXiv:2605.20761v1 Announce Type: new Abstract: The rapid proliferation of AI-generated text has introduced significant challenges in maintaining the integrity of digital content. Advanced generative

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FineVision: Open Data Is All You Need

SafetyDGX agent

arXiv:2510.17269v2 Announce Type: replace Abstract: The advancement of vision-language models (VLMs) is hampered by a fragmented landscape of inconsistent and contaminated public datasets. We introduc

FWIW, I think this moves up my AI timelines a bit. I think the next milestone will be 'Artificial *Grothendieck* Intelligence' (AGrI): defin…

IndustryDGX agent

FWIW, I think this moves up my AI timelines a bit. I think the next milestone will be 'Artificial *Grothendieck* Intelligence' (AGrI): defining new general mathematical structures to solve the hardest

Head-Aware Key-Value Compression for Efficient Autoregressive Image Generation

TutorialsDGX agent

arXiv:2605.20600v1 Announce Type: new Abstract: Autoregressive (AR) visual generation has achieved remarkable performance but suffers from high memory usage and low throughput, as it requires caching

He's right. And it has a direct implication for code review that nobody's talking about. Think about why AI cracked coding before almost eve…

AgentsDGX agent

He's right. And it has a direct implication for code review that nobody's talking about. Think about why AI cracked coding before almost everything else. Not because code is simple. Because code is ch

How Well Do Vision-Language Models Understand Sequential Driving Scenes? A Sensitivity Study

AgentsDGX agent

arXiv:2604.06750v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) are increasingly proposed for autonomous driving tasks, yet their performance on sequential driving scenes remains poo

Improving Quantized Model Performance in Qualitative Analysis with Multi-Pass Prompt Verification

Model ReleasesDGX agent

arXiv:2605.20193v1 Announce Type: new Abstract: Quantized Large Language Models (LLMs) are used more often in qualitative analysis because they run fast and need fewer computing resources. This study

Intelligent radiology workflow optimization with AI agents

ApplicationsDGX agent

Many healthcare organizations report that traditional worklist systems rely on rigid rules that ignore critical context, radiologist specialization, current workload, fatigue levels, and case complexi

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling

Model ReleasesDGX agent

arXiv:2508.08636v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized artificial intelligence by enabling complex reasoning capabilities. While recent advancements in re

JobArabi: An Arabic Corpus and Analysis of Job Announcements from Social Media

Model ReleasesDGX agent

arXiv:2605.20960v1 Announce Type: new Abstract: This paper introduces JobArabi, a large-scale corpus of Arabic job announcements collected from social media between January 2024 and October 2025. The

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

Model ReleasesDGX agent

arXiv:2605.20244v1 Announce Type: cross Abstract: We present Lean Refactor, a plug-and-play retrieval-augmented agentic framework for multi-objective, controllable, and version-robust refactoring of L

Leveraging Vision-Language Models to Detect Attention in Educational Videos

Model ReleasesDGX agent

arXiv:2605.20211v1 Announce Type: new Abstract: Educational videos are a cornerstone of remote and blended learning. However, learners' fluctuating attention remains a significant barrier to effective

Measuring and mitigating overreliance to build human-compatible AI

ApplicationsDGX agent

arXiv:2509.08010v2 Announce Type: replace-cross Abstract: Large language models (LLMs) distinguish themselves from previous technologies by functioning as collaborative ``thought partners,'' capable o

MemGym: a Long-Horizon Memory Environment for LLM Agents

Model ReleasesDGX agent

arXiv:2605.20833v1 Announce Type: new Abstract: Memory is a central capability for LLM agents operating across long-horizon tasks. Existing memory benchmarks predominantly evaluate retention of person

MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks

Model ReleasesDGX agent

arXiv:2605.20729v1 Announce Type: new Abstract: Accurate evaluation of conversational retrieval is pivotal for advancing Retrieval-Augmented Generation (RAG) systems. However, existing conversational

Multi-Week, In-Class Deployments of Telepresence Robots With Four Homebound K-12 Students: Benefits, Challenges, and Recommendations

ApplicationsDGX agent

arXiv:2605.20431v1 Announce Type: cross Abstract: Missing significant amounts of school during K-12 education is known to put students' cognitive and social development at risk. Alternatives such as h

Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches

AgentsDGX agent

arXiv:2510.04905v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have significantly improved automated code generation. While existing approaches have achieved

ShadeBench: A Benchmark Dataset for Building Shade Simulation in Sustainable Society

Model ReleasesDGX agent

arXiv:2605.20510v1 Announce Type: new Abstract: Urban heat exposure is becoming an increasingly critical challenge due to the intensifying urban heat island effect. Fine-grained shade patterns, especi

Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs

SafetyDGX agent

arXiv:2605.20191v1 Announce Type: new Abstract: Modern Large Language Models (LLMs) have recently attracted much attention for their ability to simulate human behavior and generate text that reflects

Smaller Abstract State Spaces Enable Cross-Scale Generalization in Reinforcement Learning

AgentsDGX agent

arXiv:2605.20272v1 Announce Type: new Abstract: While humans readily generalize abstract concepts to more complex or larger tasks, building Reinforcement Learning (RL) systems with this ability remain

SubTGraph: Large-Scale Subterranean Environment Synthesis with Controllable Topological Variability for Robotic Autonomy Validation

AgentsDGX agent

arXiv:2605.20917v1 Announce Type: new Abstract: Subterranean (SubT) environments have been a frontier for autonomous robotics, driven by the push for automation of mining operations and the interest i

Time-Prompt: Integrated Heterogeneous Prompts for Unlocking LLMs in Time Series Forecasting

SafetyDGX agent

arXiv:2506.17631v4 Announce Type: replace Abstract: Time series forecasting aims to model temporal dependencies among variables for future state inference, holding significant importance and widesprea

Towards UAV Detection in the Real World: A New Multispectral Dataset UAVNet-MS and a New Method

Model ReleasesDGX agent

arXiv:2605.20963v1 Announce Type: new Abstract: The proliferation of unmanned aerial vehicles (UAVs) has created urgent demand for precise UAV monitoring. Existing RGB-based systems rely on spatial cu

Training Language Agents to Learn from Experience

Model ReleasesDGX agent

arXiv:2605.20477v1 Announce Type: cross Abstract: Language agents can adapt from experience in interactive environments, but current reflection-based methods can only self-correct within a single task

Understanding and Improving Communication Performance in Multi-node LLM Inference

Model ReleasesDGX agent

arXiv:2511.09557v4 Announce Type: replace-cross Abstract: As large language models (LLMs) continue to grow in size, distributed inference has become increasingly important. Model-parallel strategies m

US government takes $2 billion equity stake in nine quantum computing firms

IndustryDGX agent

The Trump administration awarded $2 billion in grants to nine quantum-computing companies through deals that include U.S. government equity stakes , marking a departure from traditional government fun

We’re launching the Google DeepMind Accelerator program in Asia Pacific to tackle environmental risks

Model ReleasesDGX agent

Google DeepMind has launched 'AI for the Planet,' a three-month accelerator program in Asia Pacific focused on leveraging advanced AI to combat environmental challenges like climate change, biodiversi

20 May 2026

A Reproducibility Analysis of PO4ISR: Diagnosing and Mitigating Semantic Drift in LLM-Based Session Recommendation

Model ReleasesDGX agent

arXiv:2605.18780v1 Announce Type: cross Abstract: Reasoning-based Large Language Models (LLMs) like PO4ISR have set new benchmarks in session-based recommendation. However, the reproducibility of thei

Accelerating Sparse Transformer Inference on GPU

HardwareDGX agent

arXiv:2506.06095v4 Announce Type: replace Abstract: Large language models (LLMs) are popular around the world due to their powerful understanding capabilities. As the core component of LLMs, accelerat

An OpenAI model has disproved a central conjecture in discrete geometry

Model ReleasesDGX agent

An OpenAI AI model successfully disproved a longstanding conjecture in discrete geometry, a mathematical field studying geometric properties of discrete objects. This achievement demonstrates the pote

BCI-sift: An automated feature selection toolbox for Brain Computer Interface applications

TutorialsDGX agent

arXiv:2605.19646v1 Announce Type: cross Abstract: Advancements in clinical Brain-Computer Interfaces (BCIs) depend on precise and reliable signal interpretation. However, the high-dimensional and nois

BLINKG: A Benchmark for LLM-Integrated Knowledge Graph Generation

Model ReleasesDGX agent

arXiv:2605.19518v1 Announce Type: new Abstract: Generating Knowledge Graphs (KGs) remains one of the most time-consuming and labor-intensive tasks for knowledge engineers, as they need to identify sem

⚠️👇 🚨Breaking ⚠️ If we can’t make AI agents follow rules, we are screwed. New study from METR reports that “when the agents were faced wit…

SafetyDGX agent

⚠️👇 🚨Breaking ⚠️ If we can’t make AI agents follow rules, we are screwed. New study from METR reports that “when the agents were faced with hard tasks, they routinely violated constraints” This—routin

CAIT: A Syntactic Parsing Toolkit for Child-Adult InTeractions

ApplicationsDGX agent

arXiv:2605.19718v1 Announce Type: new Abstract: CHILDES is a paramount resource for language acquisition studies -- yet computational tools for analyzing its syntactic structure remain limited. Levera

CogScale: Scalable Benchmark for Sequence Processing

Model ReleasesDGX agent

arXiv:2605.19758v1 Announce Type: new Abstract: The ability to maintain and manipulate information over time is a fundamental aspect of living beings and Artificial Intelligence. While modern models h

Decentralized autonomous organization and blockchain-based incentivization framework for community-based facilities management

AgentsDGX agent

arXiv:2605.18773v1 Announce Type: cross Abstract: Traditional facility management often relies on centralized decision-making structures that limit stakeholder participation, leading to misalignment w

Discoverable Agent Knowledge -- A Formal Framework for Agentic KG Affordances (Extended Version)

AgentsDGX agent

arXiv:2605.19186v1 Announce Type: new Abstract: Two decades ago, the Semantic Web Services community was asked how agents with different ontological commitments could discover, compose, and invoke web

Dual-Gated Epistemic Time-Dilation: Autonomous Compute Modulation in Asynchronous MARL

SafetyDGX agent

arXiv:2603.23722v2 Announce Type: replace-cross Abstract: While Multi-Agent Reinforcement Learning (MARL) algorithms achieve unprecedented successes across complex continuous domains, their standard d

EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data

Model ReleasesDGX agent

arXiv:2605.19130v1 Announce Type: cross Abstract: Children acquire language grounding with remarkable robustness from limited visuo-linguistic input in ways that surpass today's best large multimodal

Federated Learning for ICD Classification with Lightweight Models and Pretrained Embeddings

Model ReleasesDGX agent

arXiv:2507.03122v2 Announce Type: replace-cross Abstract: This study investigates the feasibility and performance of federated learning (FL) for multi-label ICD code classification using clinical note

FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data

Model ReleasesDGX agent

arXiv:2605.18936v1 Announce Type: cross Abstract: Social media text data are often used to train Machine Learning (ML) models to identify users exhibiting high-risk mental health behaviors. However, s

Fine-tuning Large Language Model for Automated Algorithm Design

Model ReleasesDGX agent

arXiv:2507.10614v2 Announce Type: replace-cross Abstract: The integration of large language models (LLMs) into automated algorithm design has shown promising potential. A prevalent approach embeds LLM

FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding

Model ReleasesDGX agent

arXiv:2605.19846v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable capabilities in general video understanding, yet they often struggle with the fine-grained

Got to play with a little of this before launch as well. My experience as a social scientist was that it was more bioscience focused right n…

Model ReleasesDGX agent

Got to play with a little of this before launch as well. My experience as a social scientist was that it was more bioscience focused right now, but I think Google has been the leading lab in releasing

HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding

Model ReleasesDGX agent

arXiv:2605.19223v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) exhibit strong performance on standard video tasks, their ability to faithfully summarize and reason over

How to Model AI Agents as Personas?: Applying the Persona Ecosystem Playground to 41,300 Posts on Moltbook for Behavioral Insights

AgentsDGX agent

arXiv:2603.03140v3 Announce Type: replace-cross Abstract: AI agents are increasingly active on social media platforms, generating content and interacting with one another at scale. Yet the behavioral

Improved visual-information-driven model for crowd simulation and its modular application

SafetyDGX agent

arXiv:2504.03758v4 Announce Type: replace-cross Abstract: Crowd movement simulation is crucial for pedestrian safety management and facility design. Data-driven models offer the potential to improve r

Introducing Agent Executor, Google’s distributed Agent Runtime

Model ReleasesDGX agent

As models and harnesses improve, agents are taking on increasingly complex tasks that can run for hours or even days. But as we push agents to do more, this has surfaced a new operational problem: lon

it’s in gemini, just create it in ai studio. oh, that’s for your personal google one account. for workspace you need gemini business. no, no…

Model ReleasesDGX agent

it’s in gemini, just create it in ai studio. oh, that’s for your personal google one account. for workspace you need gemini business. no, not gemini advanced, that’s ai pro now. unless you need ai ult

K-Quantization and its Impact on Output Performance

Model ReleasesDGX agent

arXiv:2605.19645v1 Announce Type: new Abstract: Recent advancements in large language models (LLMs) have shown their remarkable capacities in many NLP tasks. However, their substantial size often pres

LLM-MC-Affect: LLM-Based Monte Carlo Modeling of Affective Trajectories and Latent Ambiguity for Interpersonal Dynamic Insight

ApplicationsDGX agent

arXiv:2601.03645v2 Announce Type: replace Abstract: Emotional coordination is a core property of human interaction that shapes how relational meaning is constructed in real time. While text-based affe

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation

Model ReleasesDGX agent

arXiv:2605.20183v1 Announce Type: new Abstract: Video generation is rapidly evolving from single-shot synthesis to complex multi-shot audio-video (MSAV) narratives to meet real-world demands. However,

NGL: Natural Garment Language for Training-Free Sewing Pattern Estimation

Model ReleasesDGX agent

arXiv:2602.20700v2 Announce Type: replace Abstract: Estimating sewing patterns from images is a practical approach for creating high-quality 3D garments, but it remains challenging due to the scarcity

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

Model ReleasesDGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

PlantTraitNet: An Uncertainty-Aware Multimodal Framework for Global-Scale Plant Trait Inference from Citizen Science Data

Model ReleasesDGX agent

arXiv:2511.06943v3 Announce Type: replace-cross Abstract: Global plant maps of plant traits, such as leaf nitrogen or plant height, are essential for understanding ecosystem processes, including the c

Position: Graph Condensation Needs a Reset -- Move Beyond Full-dataset Training and Model-Dependence

ApplicationsDGX agent

arXiv:2605.18893v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) are powerful tools for learning from graph-structured data, but their scalability is increasingly strained by the size of r

Position: Let's Develop Data Probes to Fundamentally Understand How Data Affects LLM Performance

SafetyDGX agent

arXiv:2605.18801v1 Announce Type: new Abstract: Data is fundamental to large language models (LLMs). However, understanding of what makes certain data useful for different stages of an LLM workflow, i

Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering

SafetyDGX agent

arXiv:2605.19220v1 Announce Type: cross Abstract: Uncertainty Quantification (UQ) is widely regarded as the primary safeguard for deploying Large Language Models (LLMs) in high-stakes domains. However

← Previous
1…401402403404405…428
Next →