AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,624 results
21 May 2026

Parallel LLM Reasoning for Bias-Resilient, Robust Conceptual Abstraction

SafetyDGX agent

arXiv:2605.20194v1 Announce Type: new Abstract: Large language models (LLMs) have been increasingly used to analyze text. However, they are often plagued with contextual reasoning limitations when ana

ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning

AgentsDGX agent

arXiv:2605.20342v1 Announce Type: new Abstract: Training large multimodal models (LMMs) via reinforcement learning (RL) to natively invoke video-processing tools (e.g., cropping) has become a promisin

Pareto-Enhanced Portrait Generation: Vision-Aligned Text Supervision for Alignment, Realism, and Aesthetics

SafetyDGX agent

arXiv:2605.20640v1 Announce Type: new Abstract: Text-to-image diffusion models often face a severe trilemma in human portrait generation: text-image alignment, photorealism, and human-perceived aesthe

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR

HardwareDGX agent

arXiv:2605.20863v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has recently unlocked strong reasoning capabilities in large language models (LLMs), triggering

Preference-aware Influence-function-based Data Selection Method for Efficient Fine-Tuning

SafetyDGX agent

arXiv:2605.21422v1 Announce Type: new Abstract: As LLMs continue to scale, improving training efficiency increasingly depends on using data more effectively. Data selection addresses this problem by a

proof too complicated, Claude help ELI5

Model ReleasesDGX agent

proof too complicated, Claude help ELI5 Today, we share a breakthrough on the planar unit distance problem, a famous open question first posed by Paul Erdős in 1946. For nearly 80 years, mathematician

Quantifying Hyperparameter Transfer and the Importance of Embedding Layer Learning Rate

Model ReleasesDGX agent

arXiv:2605.21486v1 Announce Type: new Abstract: Hyperparameter transfer allows extrapolating optimal optimization hyperparameters from small to large scales, making it critical for training large lang

Reducing Object Hallucination in LVLMs via Emphasizing Image-negative Tokens

TutorialsDGX agent

arXiv:2605.21300v1 Announce Type: new Abstract: Object hallucination is a significant challenge that hinders the application of large vision-language models (LVLMs) in practice. We hypothesize that on

RoadTones: Tone Controllable Text Generation from Road Event Videos

ResearchDGX agent

arXiv:2605.21411v1 Announce Type: new Abstract: Existing video-language models can generate factual descriptions of road events but lack control over how these events are expressed: their tone, urgenc

ROAR-3D: Routing Arbitrary Views for High-Fidelity 3D Generation

ResearchDGX agent

arXiv:2605.21121v1 Announce Type: new Abstract: Single-image-to-3D generative models can now produce high-quality geometry, yet conditioning on a single view inevitably introduces ambiguity about unse

Robust Personalized Recommendation under Hidden Confounding in MNAR

Model ReleasesDGX agent

arXiv:2605.21066v1 Announce Type: new Abstract: Recommender systems often rely on observational user--item interaction data, which is prone to selection bias due to users' selective interactions with

Seeing Through Fog: Towards Fog-Invariant Action Recognition

Model ReleasesDGX agent

arXiv:2605.20645v1 Announce Type: new Abstract: Foggy conditions are commonly encountered in real-world applications; however, existing action recognition approaches typically assume favorable weather

Self-Training Doesn't Flatten Language -- It Restructures It: Surface Markers Amplify While Deep Syntax Dies

ResearchDGX agent

arXiv:2605.20602v1 Announce Type: new Abstract: Successive self-training on a language model's own outputs is widely characterized as a process of flattening: diversity drops, distributions narrow, an

ShadeBench: A Benchmark Dataset for Building Shade Simulation in Sustainable Society

Model ReleasesDGX agent

arXiv:2605.20510v1 Announce Type: new Abstract: Urban heat exposure is becoming an increasingly critical challenge due to the intensifying urban heat island effect. Fine-grained shade patterns, especi

SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass

Model ReleasesDGX agent

arXiv:2602.06358v2 Announce Type: replace Abstract: We propose SHINE (Scalable Hyper In-context NEtwork), a scalable hypernetwork that can map diverse meaningful contexts into high-quality LoRA adapte

Spatial Gram Alignment for Ultra-High-Resolution Image Synthesis

SafetyDGX agent

arXiv:2605.20808v1 Announce Type: new Abstract: Modern ultra-high-resolution image synthesis relies heavily on the robust generative capacity of large-scale pre-trained Latent Diffusion Models (LDMs).

SpecBench: Measuring Reward Hacking in Long-Horizon Coding Agents

Model ReleasesDGX agent

arXiv:2605.21384v1 Announce Type: cross Abstract: As long-horizon coding agents produce more code than any developer can review, oversight collapses onto a single surface: the automated test suite. Re

Spectral Souping: A Unified Framework for Online Preference Alignment

SafetyDGX agent

arXiv:2605.20408v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) effectively aligns Large Language Models (LLMs) with aggregate human preferences but often fails to ad

Spent Tuesday watching @Google I/O thinking about a question. If you build on @Android, what just changed for you? Short answer: your screen…

Model ReleasesDGX agent

Spent Tuesday watching @Google I/O thinking about a question. If you build on @Android, what just changed for you? Short answer: your screen belongs to Gemini now. After this week, Google owns the AI,

STAR-IOD: Scale-decoupled Topology Alignment with Pseudo-label Refinement for Remote Sensing Incremental Object Detection

Model ReleasesDGX agent

arXiv:2605.20738v1 Announce Type: new Abstract: Remote sensing imagery typically arrives in the form of continuous data streams. Traditional detectors often forget previously learned categories when l

Text Analytics Evaluation Framework: A Case Study on LLMs and Social Media

Model ReleasesDGX agent

arXiv:2605.21338v1 Announce Type: new Abstract: LLMs have demonstrated exceptional proficiency in a wide range of NLP tasks. However, a notable gap remains in practical data analysis scenarios, partic

THEval. Evaluation Framework for Talking Head Video Generation

Model ReleasesDGX agent

arXiv:2511.04520v4 Announce Type: replace Abstract: Video generation has achieved remarkable progress, with generated videos increasingly resembling real ones. However, the rapid advance in generation

Time-Dependent PDE-Constrained Optimization via Weak-Form Latent Dynamics

Model ReleasesDGX agent

arXiv:2605.20639v1 Announce Type: cross Abstract: Optimization problems constrained by high-dimensional, time-dependent partial differential equations require repeated forward and sensitivity solves,

TLMs: Tiny LLMs and Agents on Edge Devices with @cormacb https://www.youtube.com/watch?v=-TiET_K-E_g Function Gemma ships at 270 million par…

Model ReleasesDGX agent

TLMs: Tiny LLMs and Agents on Edge Devices with @cormacb https://www.youtube.com/watch?v=-TiET_K-E_g Function Gemma ships at 270 million parameters and runs nearly 2,000 tokens per second prefill on a

To Select or not to Select, that is the Question: Distilling Robot Skill Prediction into a Small Ensemble

Model ReleasesDGX agent

arXiv:2605.21242v1 Announce Type: new Abstract: As robot fleets become more heterogeneous, including humanoids, rovers, quadrupeds, and drones, selecting the right robot for a task becomes a core syst

UniT: Unified Geometry Learning with Group Autoregressive Transformer

ResearchDGX agent

arXiv:2605.21131v1 Announce Type: new Abstract: Recent feed-forward models have significantly advanced geometry perception for inferring dense 3D structure from sensor observations. However, its essen

VSCD: Video-based Scene Change Detection in Unaligned Scenes

Model ReleasesDGX agent

arXiv:2605.20821v1 Announce Type: new Abstract: Detecting what has changed in an environment is essential for long-term autonomy, yet most change detection settings assume fixed viewpoints, mild misal

Weasel: Out-of-Domain Generalization for Web Agents via Importance-Diversity Data Selection

ResearchDGX agent

arXiv:2605.20291v1 Announce Type: new Abstract: Large language models (LLMs) have enabled web agents that follow natural language goals through multi-step browser interactions. However, agents fine-tu

We’re launching the Google DeepMind Accelerator program in Asia Pacific to tackle environmental risks

Model ReleasesDGX agent

Google DeepMind has launched 'AI for the Planet,' a three-month accelerator program in Asia Pacific focused on leveraging advanced AI to combat environmental challenges like climate change, biodiversi

What Twelve LLM Agent Benchmark Papers Disclose About Themselves: A Pilot Audit and an Open Scoring Schema

Model ReleasesDGX agent

arXiv:2605.21404v1 Announce Type: new Abstract: We read twelve well-known LLM agent benchmark papers and recorded, dimension by dimension, what each paper actually says about how its evaluation was ru

When to Retrain after Drift: A Data-Only Test of Post-Drift Data Size Sufficiency

Model ReleasesDGX agent

arXiv:2603.09024v2 Announce Type: replace Abstract: Sudden concept drift makes previously trained predictors unreliable, yet deciding when to retrain and what post-drift data size is sufficient is rar

Winfree Oscillatory Neural Network

Model ReleasesDGX agent

arXiv:2605.20922v1 Announce Type: cross Abstract: Oscillations and synchronization are widely believed to play a fundamental role in representation and computation. However, existing machine learning

20 May 2026

Accurate, Efficient, and Explainable Deep Learning Approaches for Environmental Science Problems

ResearchDGX agent

arXiv:2605.19366v1 Announce Type: new Abstract: Environmental science plays a pivotal role in safeguarding ecosystems, a domain driven by large-scale, heterogeneous data. In the big data era, artifici

An LLM-Based System for Argument Mining

Model ReleasesDGX agent

arXiv:2605.13793v2 Announce Type: replace Abstract: Arguments are a fundamental aspect of human reasoning, in which claims are supported, challenged, and weighed against one another. We present an end

AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration

Model ReleasesDGX agent

arXiv:2605.20025v1 Announce Type: new Abstract: Automating scientific discovery requires more than generating papers from ideas. Real research is iterative: hypotheses are challenged from multiple per

Beyond Binary Success: A Diagnostic Meta-Evaluation Framework for Fine-Grained Manipulation

Model ReleasesDGX agent

arXiv:2605.19986v1 Announce Type: cross Abstract: Fine-grained manipulation marks a regime where global scene context no longer suffices, and success hinges on the tight coupling of local attribute gr

Beyond Imitation: Learning Safe End-to-End Autonomous Driving from Hard Negatives

Model ReleasesDGX agent

arXiv:2605.19771v1 Announce Type: cross Abstract: Existing imitation learning methods for end-to-end autonomous driving predominantly learn from successful demonstrations by minimizing geometric devia

Beyond Waypoints: Dual-Heatmap Grounding for Cross-Embodiment Semantic Navigation

Model ReleasesDGX agent

arXiv:2605.19420v1 Announce Type: new Abstract: Grounding open-ended semantic instructions into physically executable local goals is a fundamental challenge in human-robot interaction. While existing

BLINKG: A Benchmark for LLM-Integrated Knowledge Graph Generation

Model ReleasesDGX agent

arXiv:2605.19518v1 Announce Type: new Abstract: Generating Knowledge Graphs (KGs) remains one of the most time-consuming and labor-intensive tasks for knowledge engineers, as they need to identify sem

Can LLMs Emulate Human Belief Dynamics?

Model ReleasesDGX agent

arXiv:2605.18781v1 Announce Type: cross Abstract: Can LLMs simulate how humans form and change beliefs in social networks? We put this to the test by replicating an established study on belief dynamic

CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization

SafetyDGX agent

arXiv:2605.19436v1 Announce Type: cross Abstract: When a model produces a correct solution under reinforcement learning with verifiable rewards (RLVR), every token receives the same reward signal rega

ClusterRAG: Cluster-Based Collaborative Filtering for Personalized Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2605.18769v1 Announce Type: cross Abstract: Personalized Retrieval-Augmented Generation (RAG) relies on accurately selecting user-relevant documents. In practice, existing RAG approaches often s

CMAD: Cooperative Multi-Agent Diffusion via Stochastic Optimal Control

AgentsDGX agent

arXiv:2602.10933v2 Announce Type: replace Abstract: Continuous-time generative models have achieved remarkable success in image restoration and synthesis. However, controlling the composition of multi

CopT: Contrastive On-Policy Thinking with Continuous Spaces for General and Agentic Reasoning

SafetyDGX agent

arXiv:2605.20075v1 Announce Type: cross Abstract: Chain-of-thought (CoT) is a standard approach for eliciting reasoning capabilities from large language models (LLMs). However, the common CoT paradigm

Critique-Guided Distillation for Robust Reasoning via Refinement

TutorialsDGX agent

arXiv:2505.11628v4 Announce Type: replace Abstract: Supervised fine-tuning with expert demonstrations often produces models that imitate outputs without internalizing the reasoning processes needed fo

Cross-Subject Intracranial EEG Reconstruction from Scalp Recordings Using Multi-Scale Cross-Attention Transformers

ResearchDGX agent

arXiv:2605.18897v1 Announce Type: cross Abstract: Intracranial EEG (iEEG) provides high-fidelity neural recordings essential for clinical and brain-computer interface applications, but acquiring these

Cross-View Splatter: Feed-Forward View Synthesis with Georeferenced Images

Model ReleasesDGX agent

arXiv:2605.19656v1 Announce Type: new Abstract: We present Cross-View Splatter, a feed-forward method that predicts pixel-aligned Gaussian splats for outdoor scenes captured at ground level AND by sat

CutVerse: A Compositional GUI Agents Benchmark for Media Post-Production Editing

Model ReleasesDGX agent

arXiv:2605.19484v1 Announce Type: cross Abstract: While GUI agents have made significant progress in web navigation and basic operating system tasks, their capabilities in professional creative workfl

D^3-Subsidy: Online and Sequential Driver Subsidy Decision-Making for Large-Scale Ride-Hailing Market

Model ReleasesDGX agent

arXiv:2605.20036v1 Announce Type: new Abstract: Ride-hailing platforms like DiDi Chuxing operate in highly dynamic environments where balancing driver supply and passenger demand is critical. Although

deadtrees.earth-aerial: A Multi-Resolution Aerial Image Dataset for Tree Cover and Mortality Detection

Model ReleasesDGX agent

arXiv:2605.19605v1 Announce Type: new Abstract: Forests worldwide are increasingly threatened by climate change and disturbances such as fire, pests, and pathogens, creating an urgent need for scalabl

Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution

TutorialsDGX agent

arXiv:2605.19228v1 Announce Type: cross Abstract: Large Language Models have achieved strong performance on reasoning tasks with objective answers by generating step-by-step solutions, but diagnosing

Does Code Cleanliness Affect Coding Agents? A Controlled Minimal-Pair Study

Model ReleasesDGX agent

arXiv:2605.20049v1 Announce Type: cross Abstract: As autonomous coding agents see rapid adoption, their evaluation has primarily focused on task completion rates holding the target codebase fixed. Thi

Efficient Conditioning Why Pseudo Observation Batch Bayesian Optimization Works When It Does not

Local AiDGX agent

arXiv:2605.18819v1 Announce Type: new Abstract: Constant Liar (CL), Kriging Believer (KB), and fantasy models are widely used for batch selection in parallel Bayesian Optimization, yet a unified theor

Efficient Pre-Training with Token Superposition

ResearchDGX agent

arXiv:2605.06546v2 Announce Type: replace Abstract: Pre-training of Large Language Models is often prohibitively expensive and inefficient at scale, requiring complex and invasive modifications in ord

EMO-BOOST: Emotion-Augmented Audio-Visual Features for Improved Generalization in Deepfake Detection

ResearchDGX agent

arXiv:2605.19630v1 Announce Type: new Abstract: With every advancement in generative AI models, forensics is under increasing pressure. The constant emergence of new generation techniques makes it imp

EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors

ResearchDGX agent

arXiv:2604.02784v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) excel at multimodal tasks, but they remain vulnerable to hallucinations that are factually incorrect or unground

EUPHORIA: Efficient Universal Planning via Hybrid Optimization for Robust Industrial Robotic Assembly

Model ReleasesDGX agent

arXiv:2605.18872v1 Announce Type: cross Abstract: Robotic assembly in architectural construction faces a persistent bottleneck: existing planners are either highly specialized, requiring prohibitive r

Evaluating the Utility of Personal Health Records in Personalized Health AI

Model ReleasesDGX agent

arXiv:2605.18937v1 Announce Type: new Abstract: Patient-managed Personal Health Records (PHRs) promises to empower patients to better understand their health; but information in the record is complex,

Explainable Wastewater Digital Twins: Adaptive Context-Conditioned Structured Simulators with Self-Falsifying Decision Support

Model ReleasesDGX agent

arXiv:2605.19826v1 Announce Type: new Abstract: Operators of safety-critical industrial processes increasingly rely on digital twins to screen control interventions, but such simulators rarely carry c

Fine-tuned a distilbert with ml-intern today for the first time. The procedure is really straightforward, almost unexpectedly so. - Found a …

Model ReleasesDGX agent

Fine-tuned a distilbert with ml-intern today for the first time. The procedure is really straightforward, almost unexpectedly so. - Found a few datasets relevant to the task (prompt injection detectio

← Previous
1…552553554555556…1061
Next →