AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
3 Jun 2026

An Exploration of Collision-based Enemy Morphology Generation

ResearchDGX agent

arXiv:2606.02832v1 Announce Type: new Abstract: Despite a great deal of prior research into Procedural Content Generation (PCG), relatively little prior work has explored generating enemies for video

Analyzing Stream Collapse in Hyper-Connections: From Diagnosis to Mitigation

ResearchDGX agent

arXiv:2606.03483v1 Announce Type: cross Abstract: Hyper-Connections (HC) replace the single Transformer residual stream with multiple streams, introducing a permutation symmetry over stream indices. W

AnchorMoE: Interpretable Time Series Classification via Anchor-Routed MoE

ApplicationsDGX agent

arXiv:2606.03631v1 Announce Type: cross Abstract: Multivariate time series classification (MTSC) is pivotal in high-stakes domains, such as clinical diagnosis and industrial fault detection, where saf


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Anomalies in Multivariate Time Series Benchmarks Are Mostly Univariate

ResearchDGX agent

arXiv:2606.02670v1 Announce Type: cross Abstract: Many recent multivariate time series anomaly detection (MT-SAD) models incorporate cross-channel modeling, under the implicit assumption that the stru

AnyAudio-Judge: A Dynamic Rubric-Based Benchmark and Evaluator for Audio Instruction Following

Model ReleasesDGX agent

arXiv:2606.03116v1 Announce Type: cross Abstract: The rapid advancement of instruction-guided audio generation has highlighted the critical need for robust alignment evaluation. Current automated eval

Approximating Probabilistic Inference in Statistical EL with Knowledge Graph Embeddings

ResearchDGX agent

arXiv:2407.11821v2 Announce Type: replace Abstract: Statistical information is ubiquitous but drawing valid conclusions from it is prohibitively hard. We explain how knowledge graph embeddings can be

Are Common Substructures Transferable? Riemannian Graph Foundation Model with Neural Vector Bundles

ResearchDGX agent

arXiv:2606.03270v1 Announce Type: cross Abstract: Foundation models have sparked a revolution via a pretraining-adaptation paradigm, with recent efforts extending this success to graphs. Unlike other

Are we really tilting? The mechanics of reward guidance in flow and diffusion models

SafetyDGX agent

arXiv:2606.02884v1 Announce Type: cross Abstract: Reward guidance algorithms steer a learned generative process toward the reward-tilted measure at inference time. While empirically powerful, these me

ASAP: Exploiting the Satisficing Generalization Edge in Neural Combinatorial Optimization

SafetyDGX agent

arXiv:2501.17377v4 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) has emerged as a promising approach for solving Combinatorial Optimization (CO) problems, such as the 3D Bin

Assessing and Mitigating Miscalibration in LLM-Based Social Science Measurement

Model ReleasesDGX agent

arXiv:2605.11954v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used in social science as scalable measurement tools for converting unstructured text into variables t

Assistax: A Multi-Agent Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics

Model ReleasesDGX agent

arXiv:2507.21638v2 Announce Type: replace Abstract: The development of reinforcement learning (RL) algorithms has been largely driven by ambitious challenge tasks and benchmarks. Games have dominated

ASymPO: Asymmetric-Scale Policy Optimization for Asynchronous LLM Post-Training Without Behavior Information

SafetyDGX agent

arXiv:2606.03070v1 Announce Type: cross Abstract: Asynchronous reinforcement learning can improve language-model post-training throughput by decoupling response generation from policy optimization, bu

Attention Calibration for Position-Fair Dense Information Retrieval

SafetyDGX agent

arXiv:2606.02737v1 Announce Type: cross Abstract: Dense retrieval models exhibit positional bias: retrieval effectiveness degrades when relevant information appears later in a passage (Zeng et al., 20

Auditable Climate Risk Intelligence from Fragmented ESG Data: Deterministic Orchestration and Imbalance-Aware Learning for Scope 1-3 Validation

Model ReleasesDGX agent

arXiv:2606.02604v1 Announce Type: cross Abstract: ESG and climate risk data remain fragmented across heterogeneous Scope 1, Scope 2, and Scope 3 reporting environments, while conventional validation p

AUDITFLOW: Executable Symbolic Environments for Structured Financial Reporting Verification

Model ReleasesDGX agent

arXiv:2606.03031v1 Announce Type: new Abstract: Structured financial audit verification is difficult for language-model agents because correctness depends on structured evidence rather than text alone

AugMask: Training Diffusion Models on Incomplete Tabular Data via Stochastic Augmentation and Masking

ApplicationsDGX agent

arXiv:2606.03347v1 Announce Type: cross Abstract: Score-based diffusion models have emerged as prominent deep generative models; however, their application to tabular data remains challenging because

AUGUSTE: Online-Learning dApp for Predictive URLLC Scheduling

AgentsDGX agent

arXiv:2606.03664v1 Announce Type: cross Abstract: Ultra Reliable and Low Latency Communications (URLLC) was one of the main motivations behind 5G, with 3GPP advertising 1-10 ms latency targets for app

AURA: Action-Gated Memory for Robot Policies at Constant VRAM

Model ReleasesDGX agent

arXiv:2606.02775v1 Announce Type: new Abstract: The KV-cache is the right memory for datacenters but the wrong memory for robots. Datacenter inference batches many short requests and resets them, amor

Automata-Conditioned Cooperative Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2511.02304v2 Announce Type: replace-cross Abstract: We study learning multi-task, multi-agent policies for cooperative, temporal objectives, under centralized training, decentralized execution.

AVTrack: Audio-Visual Tracking in Human-centric Complex Scenes

Model ReleasesDGX agent

arXiv:2606.02724v1 Announce Type: cross Abstract: Audio-visual speaker tracking aims to localize and track active speakers by leveraging auditory and visual cues, enabling fine-grained, human-centric

BAHSD: Bridging the Long-tail Gap via Adaptive Distillation in Black-box Sequential Recommendation

ResearchDGX agent

arXiv:2606.03091v1 Announce Type: cross Abstract: Sequential recommendation systems are widely adopted but often deployed as black-box APIs, which has driven recent interest in model extraction to rep

BaltiVoice: A Speech Corpus and Fine-tuned Whisper ASR System for the Balti Language

ResearchDGX agent

arXiv:2606.03504v1 Announce Type: cross Abstract: We present BaltiVoice, a 16.8-hour read-speech corpus for Balti (ISO 639-3: bft), a Tibetic language spoken in Gilgit-Baltistan, Pakistan, with no pri

BehaviorBench: Modeling Real-World User Decisions from Behavioral Traces

Model ReleasesDGX agent

arXiv:2606.02798v1 Announce Type: new Abstract: Many decision-support settings require systems that adapt to individual users, but evaluation data for this problem remain limited. Existing benchmarks

Beyond Encoder Accumulation: Measuring Encoder Roles in Multi-Encoder VLMs

Model ReleasesDGX agent

arXiv:2606.03879v1 Announce Type: cross Abstract: As foundation models scale toward fusing more heterogeneous visual streams, understanding how diverse encoders interact under joint training becomes a

BigFinanceBench: A Workflow-Grounded Benchmark for Financial-Research Agents

Model ReleasesDGX agent

arXiv:2606.03829v1 Announce Type: new Abstract: Financial-research answers are decision-relevant only when another analyst can audit how they were produced: which source was chosen, which period and a

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs

SafetyDGX agent

arXiv:2606.03647v1 Announce Type: cross Abstract: Accurately evaluating adversarial robustness is a longstanding challenge. A flawed attack design can inflate robustness estimates, making deployment r

BotDirector: Robot Storytelling Across the Symmetrical Reality with Multi-modal Interactions

AgentsDGX agent

arXiv:2606.03223v1 Announce Type: cross Abstract: Robot storytelling offers a unique blend of technological innovation and creative expression that engages children in unprecedented ways. However, the

Bridging Auxiliary Constraints to Resolve Instruction Following in Large Reasoning Models

ResearchDGX agent

arXiv:2606.03624v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated impressive capabilities in many tasks, yet they struggle with reliably following multiple instructions,

Brief Announcement: Generative Markov Model for Distributed Computing Systems

SafetyDGX agent

arXiv:2606.03061v1 Announce Type: cross Abstract: Emerging distributed computing paradigms, such as the computing continuum, are inherently heterogeneous, stochastic, and complex. Efficiently and effe

Building Better Activation Oracles

SafetyDGX agent

arXiv:2606.02609v1 Announce Type: cross Abstract: Activation Oracles (AOs) are promising methods for interpreting residual stream activations. However, current AOs face important issues, such as hallu

Building Reliable Long-Form Generation via Hallucination Rejection Sampling

ResearchDGX agent

arXiv:2606.03628v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in open-ended text generation, yet they remain prone to hallucinating incorrect or unsu

Building Trust in Black-box Optimization: A Comprehensive Framework for Explainability

ApplicationsDGX agent

arXiv:2410.14573v2 Announce Type: replace-cross Abstract: Optimizing costly black-box functions within a constrained evaluation budget presents significant challenges in many real-world applications.

Calibrating Urban Traffic Simulation from Sparse Road Observations via Genetic Optimization

ApplicationsDGX agent

arXiv:2606.03823v1 Announce Type: new Abstract: Urban traffic simulation is a critical tool for infrastructure planning, including the placement of electric vehicle charging stations. However, realist

Calibration Data Trade-offs Across Capability Dimensions: Why Multi-Source Mixing Matters for High-Sparsity LLM Pruning

Model ReleasesDGX agent

arXiv:2606.03328v1 Announce Type: cross Abstract: Post-training pruning compresses large language models to high sparsity using a small unlabelled calibration set, and recent work has concluded that t

Capability Advertisement as a Market for Lemons: A Trust Layer for Heterogeneous Agent Networks

AgentsDGX agent

arXiv:2606.03034v1 Announce Type: cross Abstract: Large language model (LLM) agents have begun to delegate work to one another. Protocols such as the Model Context Protocol (MCP) and the Agent2Agent p

CARVE: Certified Affordable Repair of Vetoed Maneuvers via Envelopes for Interactive Driving

AgentsDGX agent

arXiv:2606.02641v1 Announce Type: cross Abstract: Interactive driving exposes a failure mode that is easy to miss in rule-aware autonomous-driving stacks: a hard-rule margin can be negative for an ego

Causal Evidence of Stack Representations in Modeling Counter Languages Using Transformers

TutorialsDGX agent

arXiv:2606.03398v1 Announce Type: cross Abstract: Formal languages have proven to be effective conduits to understand the inner mechanisms of transformers. Past work has shown that transformers traine

Causal Neural Probabilistic Circuits

Model ReleasesDGX agent

arXiv:2603.01372v2 Announce Type: replace-cross Abstract: Concept Bottleneck Models (CBMs) enhance the interpretability of end-to-end neural networks by introducing a layer of concepts and predicting

Causal Preference Elicitation

Model ReleasesDGX agent

arXiv:2602.01483v2 Announce Type: replace-cross Abstract: We propose causal preference elicitation, a Bayesian framework for expert-in-the-loop causal discovery that actively queries local edge relati

CauTion: Knowing When to Trust LLMs for Ensemble Causal Discovery

ResearchDGX agent

arXiv:2606.03602v1 Announce Type: cross Abstract: Causal discovery from observational data remains challenging due to the fundamental limitations of purely statistical methods, such as statistical dis

ChatHealthAI: Aligning Electronic Health Record Representations with Large Language Models for Grounded Clinical Reasoning

Model ReleasesDGX agent

arXiv:2606.02802v1 Announce Type: new Abstract: Large language models (LLMs) exhibit strong natural-language reasoning abilities for clinical decision support, but struggle to effectively model struct

CL-DMDF:Dynamic Multimodal Data Fusion Model Based on Contrastive Learning

Local AiDGX agent

arXiv:2606.02659v1 Announce Type: cross Abstract: Multimodal data fusion involves integrating and analyzing information from multiple modalities to uncover latent correlations and complementary patter

ClinicalMC: A Benchmark for Multi-Course Clinical Decision-Making with Large Language Models

Model ReleasesDGX agent

arXiv:2606.03157v1 Announce Type: new Abstract: Large language models (LLMs) have been widely adopted in healthcare, yet they still encounter significant challenges in complex clinical decision-making

Closed-Loop Molecular Design with Calibrated Deference

AgentsDGX agent

arXiv:2606.02618v1 Announce Type: cross Abstract: We present Cognitive Loop via In-Situ Optimization (CLIO), an agent that couples a continuously-updated belief-state graph with a recursive plan-then-

Clustered Self-Assessment: A Simple yet Effective Method for Uncertainty Quantification in Large Language Models

ResearchDGX agent

arXiv:2606.03846v1 Announce Type: cross Abstract: Large language models (LLMs) demonstrate remarkable performance across diverse tasks, but they often generate responses that appear plausible while be

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization

AgentsDGX agent

arXiv:2604.17708v2 Announce Type: replace Abstract: Automating operations research (OR) with large language models (LLMs) remains limited by hand-crafted reasoning--execution workflows. Complex OR tas

Code-on-Graph: Iterative Programmatic Reasoning via Large Language Models on Knowledge Graphs

ResearchDGX agent

arXiv:2606.03705v1 Announce Type: new Abstract: Knowledge Graphs (KGs) are widely used to mitigate the limitations of Large Language Models (LLMs), such as outdated knowledge and hallucinations. Exist

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

AgentsDGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

CoEval: Ranking Language Models for Custom Tasks Without Labeled Data or Trustworthy Benchmarks

Model ReleasesDGX agent

arXiv:2606.03650v1 Announce Type: cross Abstract: Choosing or ranking language models for a specific application is hardest when no task-specific labeled data exists, and standard public benchmarks ca

Collab-REC: An LLM-based Agentic Framework for Balancing Recommendations in Tourism

SafetyDGX agent

arXiv:2508.15030v5 Announce Type: replace Abstract: We propose COLLAB-REC, a multi-agent framework designed to counteract popularity bias and improve diversity in tourism recommendations. In our setup

CoMPAS3D: A Dataset and Benchmark for Interactive Motion

Model ReleasesDGX agent

arXiv:2507.19684v2 Announce Type: replace-cross Abstract: Socially interactive humanoid robots must engage with humans through their bodies, adapting in real time to a partner's movement, intent, and

Conditional Hypothesis Generation for LLM-Based Text Analysis with Researcher-Specified Covariates

ApplicationsDGX agent

arXiv:2606.03029v1 Announce Type: cross Abstract: A core goal of computational social science is to discover interpretable differences in how language varies across outcomes of interest, such as polit

Conditional Latent Diffusion Model with Fourier-based Motion Modelling for Virtual Population Synthesis

ResearchDGX agent

arXiv:2606.03827v1 Announce Type: cross Abstract: In-silico trials of medical devices require the generation of virtual populations of anatomies. In cardiovascular applications, virtual anatomy is typ

Consistency Training Can Entrench Misalignment

SafetyDGX agent

arXiv:2606.03810v1 Announce Type: cross Abstract: Consistency training encourages a model to produce similar outputs across related inputs or sampling procedures. Such methods are simple, scalable, an

Constitutional On-Policy Safe Distillation

SafetyDGX agent

arXiv:2606.03089v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) has emerged as an efficient post-training paradigm by using a teacher conditioned on privileged information to prov

ConTraIRL: Factorized Contrastive Abstractions for Transferable IRL

SafetyDGX agent

arXiv:2606.03017v1 Announce Type: cross Abstract: Reward transfer in Inverse Reinforcement Learning (IRL) is unreliable when policies must generalize to unseen combinations of environment dynamics and

CORE: Conflict-Oriented Reasoning for General Multimodal Manipulation Detection

ResearchDGX agent

arXiv:2606.03066v1 Announce Type: new Abstract: The rapid rise of generative AI has made multimodal fake news increasingly realistic and pervasive, posing severe threats to public trust and social sta

Cosmos 3: Omnimodal World Models for Physical AI

Model ReleasesDGX agent

arXiv:2606.02800v1 Announce Type: cross Abstract: We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences

Cost-Aware Query Routing in RAG: Empirical Analysis of Retrieval Depth Tradeoffs

Model ReleasesDGX agent

arXiv:2606.02581v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) faces a fundamental three-way tension: deeper retrieval improves factual grounding but inflates token costs and e

Coupled Local and Global World Models for Efficient First Order RL

Local AiDGX agent

arXiv:2602.06219v2 Announce Type: replace-cross Abstract: World models offer a promising avenue for more faithfully capturing complex dynamics, including contacts and non-rigidity, as well as complex

← Previous
1…163164165166167…358
Next →