AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
3 Jun 2026

Can LLM Rerankers Predict Their Own Ranking Performance?

ResearchDGX agent

arXiv:2606.03535v1 Announce Type: cross Abstract: Retrieval effectiveness varies substantially across queries, making it important to estimate ranking quality before relevance judgments are available.

Can Local Learning Match Self-Supervised Backpropagation?

ResearchDGX agent

arXiv:2601.21683v2 Announce Type: replace Abstract: While end-to-end self-supervised learning with backpropagation (global BP-SSL) has become central for training modern AI systems, theories of local

Can Structural Cues Save LLMs? Evaluating Language Models in Massive Document Streams

Model ReleasesDGX agent

arXiv:2603.19250v2 Announce Type: replace Abstract: Evaluating language models in streaming environments is critical, yet underexplored. Existing benchmarks either focus on single complex events or pr

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CANMOT: Class-Aware Noise Modeling for Multi-Object Tracking in Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.03590v1 Announce Type: new Abstract: Kalman filter (KF)-based multi-object tracking (MOT) remains a strong baseline for autonomous driving due to its strong performance, computational effic

Capability Advertisement as a Market for Lemons: A Trust Layer for Heterogeneous Agent Networks

AgentsDGX agent

arXiv:2606.03034v1 Announce Type: cross Abstract: Large language model (LLM) agents have begun to delegate work to one another. Protocols such as the Model Context Protocol (MCP) and the Agent2Agent p

CAPER: Clause-Aligned Process Supervision for Text-to-SQL

Model ReleasesDGX agent

arXiv:2606.03327v1 Announce Type: cross Abstract: Text-to-SQL systems are typically evaluated by query-level execution correctness, but this terminal signal provides little guidance about which interm

CARVE: Certified Affordable Repair of Vetoed Maneuvers via Envelopes for Interactive Driving

AgentsDGX agent

arXiv:2606.02641v1 Announce Type: cross Abstract: Interactive driving exposes a failure mode that is easy to miss in rule-aware autonomous-driving stacks: a hard-rule margin can be negative for an ego

Causal Evidence of Stack Representations in Modeling Counter Languages Using Transformers

TutorialsDGX agent

arXiv:2606.03398v1 Announce Type: cross Abstract: Formal languages have proven to be effective conduits to understand the inner mechanisms of transformers. Past work has shown that transformers traine

Causal Neural Probabilistic Circuits

Model ReleasesDGX agent

arXiv:2603.01372v2 Announce Type: replace-cross Abstract: Concept Bottleneck Models (CBMs) enhance the interpretability of end-to-end neural networks by introducing a layer of concepts and predicting

Causal Preference Elicitation

Model ReleasesDGX agent

arXiv:2602.01483v2 Announce Type: replace-cross Abstract: We propose causal preference elicitation, a Bayesian framework for expert-in-the-loop causal discovery that actively queries local edge relati

CauTion: Knowing When to Trust LLMs for Ensemble Causal Discovery

ResearchDGX agent

arXiv:2606.03602v1 Announce Type: cross Abstract: Causal discovery from observational data remains challenging due to the fundamental limitations of purely statistical methods, such as statistical dis

Characterizing Detectability in 3DGS Poisoning: A Stage-wise Benchmark

Model ReleasesDGX agent

arXiv:2606.03499v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has rapidly emerged as a leading representation for real-time novel view synthesis, but recent work shows it is vulnerable

Chatbots Output Meaningful (but Problematic) Language

Model ReleasesDGX agent

arXiv:2606.02973v1 Announce Type: new Abstract: Are utterances by AI chatbots meaningful? Concretely, if a user asks, say, Anthropic's agent Claude, 'What is the capital of Spain?' and Claude answers,

ChatHealthAI: Aligning Electronic Health Record Representations with Large Language Models for Grounded Clinical Reasoning

Model ReleasesDGX agent

arXiv:2606.02802v1 Announce Type: new Abstract: Large language models (LLMs) exhibit strong natural-language reasoning abilities for clinical decision support, but struggle to effectively model struct

CL-DMDF:Dynamic Multimodal Data Fusion Model Based on Contrastive Learning

Local AiDGX agent

arXiv:2606.02659v1 Announce Type: cross Abstract: Multimodal data fusion involves integrating and analyzing information from multiple modalities to uncover latent correlations and complementary patter

ClinicalMC: A Benchmark for Multi-Course Clinical Decision-Making with Large Language Models

Model ReleasesDGX agent

arXiv:2606.03157v1 Announce Type: new Abstract: Large language models (LLMs) have been widely adopted in healthcare, yet they still encounter significant challenges in complex clinical decision-making

Closed-Loop Molecular Design with Calibrated Deference

AgentsDGX agent

arXiv:2606.02618v1 Announce Type: cross Abstract: We present Cognitive Loop via In-Situ Optimization (CLIO), an agent that couples a continuously-updated belief-state graph with a recursive plan-then-

Clustered Self-Assessment: A Simple yet Effective Method for Uncertainty Quantification in Large Language Models

ResearchDGX agent

arXiv:2606.03846v1 Announce Type: cross Abstract: Large language models (LLMs) demonstrate remarkable performance across diverse tasks, but they often generate responses that appear plausible while be

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization

AgentsDGX agent

arXiv:2604.17708v2 Announce Type: replace Abstract: Automating operations research (OR) with large language models (LLMs) remains limited by hand-crafted reasoning--execution workflows. Complex OR tas

COD10K-C: Benchmarking Robustness of Camouflaged Object Detection Under Natural Image Corruptions

Model ReleasesDGX agent

arXiv:2606.02603v1 Announce Type: new Abstract: Camouflaged object detection has improved substantially, but most standard benchmarks evaluate models only on clean images. This is not realistic becaus

Code-on-Graph: Iterative Programmatic Reasoning via Large Language Models on Knowledge Graphs

ResearchDGX agent

arXiv:2606.03705v1 Announce Type: new Abstract: Knowledge Graphs (KGs) are widely used to mitigate the limitations of Large Language Models (LLMs), such as outdated knowledge and hallucinations. Exist

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

AgentsDGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

CoEval: Ranking Language Models for Custom Tasks Without Labeled Data or Trustworthy Benchmarks

Model ReleasesDGX agent

arXiv:2606.03650v1 Announce Type: cross Abstract: Choosing or ranking language models for a specific application is hardest when no task-specific labeled data exists, and standard public benchmarks ca

Coherence Maximization Improves Pluralistic Alignment

SafetyDGX agent

arXiv:2606.03110v1 Announce Type: new Abstract: Aligning AI systems with diverse human values requires value specifications grounded in concrete examples, but generating such examples without extensiv

Coherent Swap Regret and Channel-Proof Learning

Model ReleasesDGX agent

arXiv:2606.02655v1 Announce Type: cross Abstract: External regret certifies stability only against replacing one's behavior by a fixed alternative. In a quantum game, this misses a natural physical mo

Collab-REC: An LLM-based Agentic Framework for Balancing Recommendations in Tourism

SafetyDGX agent

arXiv:2508.15030v5 Announce Type: replace Abstract: We propose COLLAB-REC, a multi-agent framework designed to counteract popularity bias and improve diversity in tourism recommendations. In our setup

Combining Statistical Features and Deep Encodings for Rehearsal-Based Class-Incremental Time Series Classification

Model ReleasesDGX agent

arXiv:2606.03292v1 Announce Type: cross Abstract: Many systems used in real-world environments require adding new categories and incorporating new information without forgetting what was previously le

CoMPAS3D: A Dataset and Benchmark for Interactive Motion

Model ReleasesDGX agent

arXiv:2507.19684v2 Announce Type: replace-cross Abstract: Socially interactive humanoid robots must engage with humans through their bodies, adapting in real time to a partner's movement, intent, and

Compress then Merge: From Multiple LoRAs into One Low-Rank Adapter

Model ReleasesDGX agent

arXiv:2606.03723v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) enables parameter-efficient specialization of foundation models, but the proliferation of task-specific adapters fragments ca

Conditional Hypothesis Generation for LLM-Based Text Analysis with Researcher-Specified Covariates

ApplicationsDGX agent

arXiv:2606.03029v1 Announce Type: cross Abstract: A core goal of computational social science is to discover interpretable differences in how language varies across outcomes of interest, such as polit

Conditional Latent Diffusion Model with Fourier-based Motion Modelling for Virtual Population Synthesis

ResearchDGX agent

arXiv:2606.03827v1 Announce Type: cross Abstract: In-silico trials of medical devices require the generation of virtual populations of anatomies. In cardiovascular applications, virtual anatomy is typ

Conformal Language Modeling via Posterior Sampling

ResearchDGX agent

arXiv:2606.03731v1 Announce Type: new Abstract: Large Language Models remain plagued by hallucinations. Recent work has sought to tame their prevalence using statistical techniques based on conformal

Consistency Training Can Entrench Misalignment

SafetyDGX agent

arXiv:2606.03810v1 Announce Type: cross Abstract: Consistency training encourages a model to produce similar outputs across related inputs or sampling procedures. Such methods are simple, scalable, an

Consistent Yet Wrong: Evidence Insensitivity in Spatial Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.02742v1 Announce Type: new Abstract: Spatial reasoning is fundamental to robotics, autonomy, and embodied AI, yet modern vision-language models (VLMs) remain unreliable on metric distance q

Constitutional On-Policy Safe Distillation

SafetyDGX agent

arXiv:2606.03089v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) has emerged as an efficient post-training paradigm by using a teacher conditioned on privileged information to prov

ConTrack: Constrained Hand Motion Tracking with Adaptive Trade-off Control

SafetyDGX agent

arXiv:2606.03177v1 Announce Type: new Abstract: Human demonstrations provide strong priors for robot manipulation, yet it is non-trivial to transfer them to execute on real robots due to the kinematic

ConTraIRL: Factorized Contrastive Abstractions for Transferable IRL

SafetyDGX agent

arXiv:2606.03017v1 Announce Type: cross Abstract: Reward transfer in Inverse Reinforcement Learning (IRL) is unreliable when policies must generalize to unseen combinations of environment dynamics and

Contrastive Neural Algorithmic Reasoning for Graph Coloring

SafetyDGX agent

arXiv:2606.03923v1 Announce Type: new Abstract: Graph coloring seeks to assigns colors to a graph's nodes so that adjacent nodes receive different colors, using as few colors as possible. Here, we stu

CoralBay: A Self-Supervised CT Foundation Model

Model ReleasesDGX agent

arXiv:2606.03888v1 Announce Type: new Abstract: Self-supervised learning has enabled large-scale pre-training on 2D natural images, producing general-purpose visual representations that transfer effec

Core-based Hierarchies for Efficient GraphRAG

ApplicationsDGX agent

arXiv:2603.05207v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) enhances large language models by incorporating external knowledge. However, existing vector-based method

CORE: Conflict-Oriented Reasoning for General Multimodal Manipulation Detection

ResearchDGX agent

arXiv:2606.03066v1 Announce Type: new Abstract: The rapid rise of generative AI has made multimodal fake news increasingly realistic and pervasive, posing severe threats to public trust and social sta

Correcting Neural Operator Spectral Bias via Diffusion Posterior Sampling with Sparse Observations

SafetyDGX agent

arXiv:2606.03936v1 Announce Type: new Abstract: Neural operator surrogates (NO) approximate PDE solutions orders of magnitude faster than numerical solvers, but suffer from spectral bias: high-frequen

Cosmos 3: Omnimodal World Models for Physical AI

Model ReleasesDGX agent

arXiv:2606.02800v1 Announce Type: cross Abstract: We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences

Cost-Aware Query Routing in RAG: Empirical Analysis of Retrieval Depth Tradeoffs

Model ReleasesDGX agent

arXiv:2606.02581v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) faces a fundamental three-way tension: deeper retrieval improves factual grounding but inflates token costs and e

CoughSense: Five-Class Respiratory Disease Classification via Whisper Encoder Fine-Tuning and Dual-Encoder Cross-Attention Fusion with Balanced Contrastive Learning

ResearchDGX agent

arXiv:2606.02998v1 Announce Type: new Abstract: Automated cough analysis offers a path to low-cost respiratory screening, but most existing work stops at binary COVID-19 detection. A practical tool ne

Coupled Local and Global World Models for Efficient First Order RL

Local AiDGX agent

arXiv:2602.06219v2 Announce Type: replace-cross Abstract: World models offer a promising avenue for more faithfully capturing complex dynamics, including contacts and non-rigidity, as well as complex

CourseTimeQA: A Lecture-Video Benchmark and a Latency-Constrained Cross-Modal Fusion Method for Timestamped QA

Model ReleasesDGX agent

arXiv:2512.00360v2 Announce Type: replace Abstract: We study timestamped question answering over educational lecture videos under a single-GPU latency/memory budget. Given a natural-language query, th

CP-Agent: Context-Aware Multimodal Reasoning for Cellular Morphological Profiling under Chemical Perturbations

SafetyDGX agent

arXiv:2606.03435v1 Announce Type: new Abstract: Cell Painting combines multiplexed fluorescent staining, high-content imaging, and quantitative analysis to generate high-dimensional phenotypic readout

CRAM-ER: Error-Resilient Spintronic Computational Random Access Memory for Scalable In-Memory Computation

HardwareDGX agent

arXiv:2606.02781v1 Announce Type: cross Abstract: Deep neural networks (DNNs) have achieved state-of-the-art performance across diverse domains. However, typical Von Neumann compute paradigms face sev

CREward: A Type-Specific Creativity Reward Model

Model ReleasesDGX agent

arXiv:2511.19995v2 Announce Type: replace Abstract: Creativity is a complex phenomenon. When it comes to representing and assessing creativity, treating it as a single undifferentiated quantity would

Critical evaluation of PINN for FWD inverse analysis and differentiable FEM as an alternative

Model ReleasesDGX agent

arXiv:2606.03210v1 Announce Type: cross Abstract: Automatic-differentiation-based inverse analysis methods, including physics-informed neural networks (PINNs) and differentiable programming, have rece

CropCraft: A Procedural World Generator for Robotic Simulation of Agricultural Tasks

ResearchDGX agent

arXiv:2511.02417v2 Announce Type: replace Abstract: The adoption of agroecological practices in modern agriculture requires robotic systems capable of operating in highly diverse and complex field env

Cross-Lingual Token Arbitrage: Optimizing Code Agent Context Windows via Local LLM Preprocessing

Model ReleasesDGX agent

arXiv:2606.03618v1 Announce Type: new Abstract: AI-assisted coding agents are bottlenecked by input-token cost. Two pathologies of raw human input drive much of this overhead: tokenization inefficienc

Cross-Modal Contrastive Learning of ECG and Angiography Representations for Severe Stenosis Classification

ResearchDGX agent

arXiv:2606.02605v1 Announce Type: cross Abstract: Coronary artery stenosis is a common cardiovascular disease, with severe, untreated cases posing significant risks of heart attack. Although coronary

Cross-Modality Feature Fusion Based on Structured State Space Duality for Multimodal Image Registration Network

Local AiDGX agent

arXiv:2606.03341v1 Announce Type: new Abstract: In multi-modal image registration, the primary challenge lies in shared structural information extraction. Compared to Transformers, Structured State Sp

Cryo-Bench: Benchmarking Foundation Models for Cryosphere Applications

Model ReleasesDGX agent

arXiv:2603.01576v3 Announce Type: replace Abstract: Geo-Foundation Models (GFMs) have been evaluated across diverse Earth observation task including multiple domains and have demonstrated strong poten

CTR-Sink: Attention Sink for Language Models in Click-Through Rate Prediction

ResearchDGX agent

arXiv:2508.03668v2 Announce Type: replace Abstract: Click-Through Rate (CTR) prediction, a core task in recommendation systems, estimates user click likelihood using historical behavioral data. Modeli

Curriculum-Adapted Robust Reinforcement Learning for UAV Deconfliction in Adversarial Environments

SafetyDGX agent

arXiv:2506.21129v2 Announce Type: replace-cross Abstract: Autonomous unmanned aerial vehicles (UAVs) increasingly rely on reinforcement learning (RL) for navigation. However, global navigation satelli

D-Judge: Disrupting Multi-Turn Jailbreaks using Semantics-Preserving Output Rewriting

SafetyDGX agent

arXiv:2606.02640v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to large language model (LLM) safety because they exploit feedback from auxiliary judge models to i

Data- and Variance-dependent Regret Bounds for Online Tabular MDPs

SafetyDGX agent

arXiv:2602.01903v2 Announce Type: replace Abstract: This work studies online episodic tabular Markov decision processes (MDPs) with known transitions and develops best-of-both-worlds algorithms that a

← Previous
1…492493494495496…1049
Next →