AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlog
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
22,241 results
Research

Analyzing Stream Collapse in Hyper-Connections: From Diagnosis to Mitigation

DGX agent

arXiv:2606.03483v1 Announce Type: cross Abstract: Hyper-Connections (HC) replace the single Transformer residual stream with multiple streams, introducing a permutation symmetry over stream indices. W

researcharxiv-cs-ai
3 Jun 2026
Applications
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

AnchorMoE: Interpretable Time Series Classification via Anchor-Routed MoE

DGX agent

arXiv:2606.03631v1 Announce Type: cross Abstract: Multivariate time series classification (MTSC) is pivotal in high-stakes domains, such as clinical diagnosis and industrial fault detection, where saf

applicationsarxiv-cs-ai
3 Jun 2026
Research

Anomalies in Multivariate Time Series Benchmarks Are Mostly Univariate

DGX agent

arXiv:2606.02670v1 Announce Type: cross Abstract: Many recent multivariate time series anomaly detection (MT-SAD) models incorporate cross-channel modeling, under the implicit assumption that the stru

researcharxiv-cs-ai
3 Jun 2026
Model Releases

AnyAudio-Judge: A Dynamic Rubric-Based Benchmark and Evaluator for Audio Instruction Following

DGX agent

arXiv:2606.03116v1 Announce Type: cross Abstract: The rapid advancement of instruction-guided audio generation has highlighted the critical need for robust alignment evaluation. Current automated eval

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Approximating Probabilistic Inference in Statistical EL with Knowledge Graph Embeddings

DGX agent

arXiv:2407.11821v2 Announce Type: replace Abstract: Statistical information is ubiquitous but drawing valid conclusions from it is prohibitively hard. We explain how knowledge graph embeddings can be

researcharxiv-cs-ai
3 Jun 2026
Research

Are Common Substructures Transferable? Riemannian Graph Foundation Model with Neural Vector Bundles

DGX agent

arXiv:2606.03270v1 Announce Type: cross Abstract: Foundation models have sparked a revolution via a pretraining-adaptation paradigm, with recent efforts extending this success to graphs. Unlike other

researcharxiv-cs-ai
3 Jun 2026
Safety

Are we really tilting? The mechanics of reward guidance in flow and diffusion models

DGX agent

arXiv:2606.02884v1 Announce Type: cross Abstract: Reward guidance algorithms steer a learned generative process toward the reward-tilted measure at inference time. While empirically powerful, these me

safetyarxiv-cs-ai
3 Jun 2026
Safety

ASAP: Exploiting the Satisficing Generalization Edge in Neural Combinatorial Optimization

DGX agent

arXiv:2501.17377v4 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) has emerged as a promising approach for solving Combinatorial Optimization (CO) problems, such as the 3D Bin

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Assessing and Mitigating Miscalibration in LLM-Based Social Science Measurement

DGX agent

arXiv:2605.11954v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used in social science as scalable measurement tools for converting unstructured text into variables t

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Assistax: A Multi-Agent Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics

DGX agent

arXiv:2507.21638v2 Announce Type: replace Abstract: The development of reinforcement learning (RL) algorithms has been largely driven by ambitious challenge tasks and benchmarks. Games have dominated

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

ASymPO: Asymmetric-Scale Policy Optimization for Asynchronous LLM Post-Training Without Behavior Information

DGX agent

arXiv:2606.03070v1 Announce Type: cross Abstract: Asynchronous reinforcement learning can improve language-model post-training throughput by decoupling response generation from policy optimization, bu

safetyarxiv-cs-ai
3 Jun 2026
Safety

Attention Calibration for Position-Fair Dense Information Retrieval

DGX agent

arXiv:2606.02737v1 Announce Type: cross Abstract: Dense retrieval models exhibit positional bias: retrieval effectiveness degrades when relevant information appears later in a passage (Zeng et al., 20

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Auditable Climate Risk Intelligence from Fragmented ESG Data: Deterministic Orchestration and Imbalance-Aware Learning for Scope 1-3 Validation

DGX agent

arXiv:2606.02604v1 Announce Type: cross Abstract: ESG and climate risk data remain fragmented across heterogeneous Scope 1, Scope 2, and Scope 3 reporting environments, while conventional validation p

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

AUDITFLOW: Executable Symbolic Environments for Structured Financial Reporting Verification

DGX agent

arXiv:2606.03031v1 Announce Type: new Abstract: Structured financial audit verification is difficult for language-model agents because correctness depends on structured evidence rather than text alone

model-releasesarxiv-cs-ai
3 Jun 2026
Applications

AugMask: Training Diffusion Models on Incomplete Tabular Data via Stochastic Augmentation and Masking

DGX agent

arXiv:2606.03347v1 Announce Type: cross Abstract: Score-based diffusion models have emerged as prominent deep generative models; however, their application to tabular data remains challenging because

applicationsarxiv-cs-ai
3 Jun 2026
Agents

AUGUSTE: Online-Learning dApp for Predictive URLLC Scheduling

DGX agent

arXiv:2606.03664v1 Announce Type: cross Abstract: Ultra Reliable and Low Latency Communications (URLLC) was one of the main motivations behind 5G, with 3GPP advertising 1-10 ms latency targets for app

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

AURA: Action-Gated Memory for Robot Policies at Constant VRAM

DGX agent

arXiv:2606.02775v1 Announce Type: new Abstract: The KV-cache is the right memory for datacenters but the wrong memory for robots. Datacenter inference batches many short requests and resets them, amor

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Automata-Conditioned Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2511.02304v2 Announce Type: replace-cross Abstract: We study learning multi-task, multi-agent policies for cooperative, temporal objectives, under centralized training, decentralized execution.

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

AVTrack: Audio-Visual Tracking in Human-centric Complex Scenes

DGX agent

arXiv:2606.02724v1 Announce Type: cross Abstract: Audio-visual speaker tracking aims to localize and track active speakers by leveraging auditory and visual cues, enabling fine-grained, human-centric

model-releasesarxiv-cs-ai
3 Jun 2026
Research

BAHSD: Bridging the Long-tail Gap via Adaptive Distillation in Black-box Sequential Recommendation

DGX agent

arXiv:2606.03091v1 Announce Type: cross Abstract: Sequential recommendation systems are widely adopted but often deployed as black-box APIs, which has driven recent interest in model extraction to rep

researcharxiv-cs-ai
3 Jun 2026
Research

BaltiVoice: A Speech Corpus and Fine-tuned Whisper ASR System for the Balti Language

DGX agent

arXiv:2606.03504v1 Announce Type: cross Abstract: We present BaltiVoice, a 16.8-hour read-speech corpus for Balti (ISO 639-3: bft), a Tibetic language spoken in Gilgit-Baltistan, Pakistan, with no pri

researcharxiv-cs-ai
3 Jun 2026
Model Releases

BehaviorBench: Modeling Real-World User Decisions from Behavioral Traces

DGX agent

arXiv:2606.02798v1 Announce Type: new Abstract: Many decision-support settings require systems that adapt to individual users, but evaluation data for this problem remain limited. Existing benchmarks

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Beyond Encoder Accumulation: Measuring Encoder Roles in Multi-Encoder VLMs

DGX agent

arXiv:2606.03879v1 Announce Type: cross Abstract: As foundation models scale toward fusing more heterogeneous visual streams, understanding how diverse encoders interact under joint training becomes a

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

BigFinanceBench: A Workflow-Grounded Benchmark for Financial-Research Agents

DGX agent

arXiv:2606.03829v1 Announce Type: new Abstract: Financial-research answers are decision-relevant only when another analyst can audit how they were produced: which source was chosen, which period and a

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs

DGX agent

arXiv:2606.03647v1 Announce Type: cross Abstract: Accurately evaluating adversarial robustness is a longstanding challenge. A flawed attack design can inflate robustness estimates, making deployment r

safetyarxiv-cs-ai
3 Jun 2026
Agents

BotDirector: Robot Storytelling Across the Symmetrical Reality with Multi-modal Interactions

DGX agent

arXiv:2606.03223v1 Announce Type: cross Abstract: Robot storytelling offers a unique blend of technological innovation and creative expression that engages children in unprecedented ways. However, the

agentsarxiv-cs-ai
3 Jun 2026
Research

Bridging Auxiliary Constraints to Resolve Instruction Following in Large Reasoning Models

DGX agent

arXiv:2606.03624v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated impressive capabilities in many tasks, yet they struggle with reliably following multiple instructions,

researcharxiv-cs-ai
3 Jun 2026
Safety

Brief Announcement: Generative Markov Model for Distributed Computing Systems

DGX agent

arXiv:2606.03061v1 Announce Type: cross Abstract: Emerging distributed computing paradigms, such as the computing continuum, are inherently heterogeneous, stochastic, and complex. Efficiently and effe

safetyarxiv-cs-ai
3 Jun 2026
Safety

Building Better Activation Oracles

DGX agent

arXiv:2606.02609v1 Announce Type: cross Abstract: Activation Oracles (AOs) are promising methods for interpreting residual stream activations. However, current AOs face important issues, such as hallu

safetyarxiv-cs-ai
3 Jun 2026
Research

Building Reliable Long-Form Generation via Hallucination Rejection Sampling

DGX agent

arXiv:2606.03628v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in open-ended text generation, yet they remain prone to hallucinating incorrect or unsu

researcharxiv-cs-ai
3 Jun 2026
Applications

Building Trust in Black-box Optimization: A Comprehensive Framework for Explainability

DGX agent

arXiv:2410.14573v2 Announce Type: replace-cross Abstract: Optimizing costly black-box functions within a constrained evaluation budget presents significant challenges in many real-world applications.

applicationsarxiv-cs-ai
3 Jun 2026
Applications

Calibrating Urban Traffic Simulation from Sparse Road Observations via Genetic Optimization

DGX agent

arXiv:2606.03823v1 Announce Type: new Abstract: Urban traffic simulation is a critical tool for infrastructure planning, including the placement of electric vehicle charging stations. However, realist

applicationsarxiv-cs-ai
3 Jun 2026
Model Releases

Calibration Data Trade-offs Across Capability Dimensions: Why Multi-Source Mixing Matters for High-Sparsity LLM Pruning

DGX agent

arXiv:2606.03328v1 Announce Type: cross Abstract: Post-training pruning compresses large language models to high sparsity using a small unlabelled calibration set, and recent work has concluded that t

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Capability Advertisement as a Market for Lemons: A Trust Layer for Heterogeneous Agent Networks

DGX agent

arXiv:2606.03034v1 Announce Type: cross Abstract: Large language model (LLM) agents have begun to delegate work to one another. Protocols such as the Model Context Protocol (MCP) and the Agent2Agent p

agentsarxiv-cs-ai
3 Jun 2026
Agents

CARVE: Certified Affordable Repair of Vetoed Maneuvers via Envelopes for Interactive Driving

DGX agent

arXiv:2606.02641v1 Announce Type: cross Abstract: Interactive driving exposes a failure mode that is easy to miss in rule-aware autonomous-driving stacks: a hard-rule margin can be negative for an ego

agentsarxiv-cs-ai
3 Jun 2026
Tutorials

Causal Evidence of Stack Representations in Modeling Counter Languages Using Transformers

DGX agent

arXiv:2606.03398v1 Announce Type: cross Abstract: Formal languages have proven to be effective conduits to understand the inner mechanisms of transformers. Past work has shown that transformers traine

tutorialsarxiv-cs-ai
3 Jun 2026
Model Releases

Causal Neural Probabilistic Circuits

DGX agent

arXiv:2603.01372v2 Announce Type: replace-cross Abstract: Concept Bottleneck Models (CBMs) enhance the interpretability of end-to-end neural networks by introducing a layer of concepts and predicting

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Causal Preference Elicitation

DGX agent

arXiv:2602.01483v2 Announce Type: replace-cross Abstract: We propose causal preference elicitation, a Bayesian framework for expert-in-the-loop causal discovery that actively queries local edge relati

model-releasesarxiv-cs-ai
3 Jun 2026
Research

CauTion: Knowing When to Trust LLMs for Ensemble Causal Discovery

DGX agent

arXiv:2606.03602v1 Announce Type: cross Abstract: Causal discovery from observational data remains challenging due to the fundamental limitations of purely statistical methods, such as statistical dis

researcharxiv-cs-ai
3 Jun 2026
Model Releases

ChatHealthAI: Aligning Electronic Health Record Representations with Large Language Models for Grounded Clinical Reasoning

DGX agent

arXiv:2606.02802v1 Announce Type: new Abstract: Large language models (LLMs) exhibit strong natural-language reasoning abilities for clinical decision support, but struggle to effectively model struct

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

CL-DMDF:Dynamic Multimodal Data Fusion Model Based on Contrastive Learning

DGX agent

arXiv:2606.02659v1 Announce Type: cross Abstract: Multimodal data fusion involves integrating and analyzing information from multiple modalities to uncover latent correlations and complementary patter

local-aiarxiv-cs-ai
3 Jun 2026
Model Releases

ClinicalMC: A Benchmark for Multi-Course Clinical Decision-Making with Large Language Models

DGX agent

arXiv:2606.03157v1 Announce Type: new Abstract: Large language models (LLMs) have been widely adopted in healthcare, yet they still encounter significant challenges in complex clinical decision-making

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Closed-Loop Molecular Design with Calibrated Deference

DGX agent

arXiv:2606.02618v1 Announce Type: cross Abstract: We present Cognitive Loop via In-Situ Optimization (CLIO), an agent that couples a continuously-updated belief-state graph with a recursive plan-then-

agentsarxiv-cs-ai
3 Jun 2026
Research

Clustered Self-Assessment: A Simple yet Effective Method for Uncertainty Quantification in Large Language Models

DGX agent

arXiv:2606.03846v1 Announce Type: cross Abstract: Large language models (LLMs) demonstrate remarkable performance across diverse tasks, but they often generate responses that appear plausible while be

researcharxiv-cs-ai
3 Jun 2026
Agents

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization

DGX agent

arXiv:2604.17708v2 Announce Type: replace Abstract: Automating operations research (OR) with large language models (LLMs) remains limited by hand-crafted reasoning--execution workflows. Complex OR tas

agentsarxiv-cs-ai
3 Jun 2026
Research

Code-on-Graph: Iterative Programmatic Reasoning via Large Language Models on Knowledge Graphs

DGX agent

arXiv:2606.03705v1 Announce Type: new Abstract: Knowledge Graphs (KGs) are widely used to mitigate the limitations of Large Language Models (LLMs), such as outdated knowledge and hallucinations. Exist

researcharxiv-cs-ai
3 Jun 2026
Agents

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

DGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

CoEval: Ranking Language Models for Custom Tasks Without Labeled Data or Trustworthy Benchmarks

DGX agent

arXiv:2606.03650v1 Announce Type: cross Abstract: Choosing or ranking language models for a specific application is hardest when no task-specific labeled data exists, and standard public benchmarks ca

model-releasesarxiv-cs-ai
3 Jun 2026
← Previous
1…220221222223224…464
Next →