AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
9 Jun 2026

BareWave: Waveform-Native Flow-Matching Text-to-Speech

SafetyDGX agent

arXiv:2606.09048v1 Announce Type: cross Abstract: Removing intermediate representations and separately trained decoding stages has become an important direction in generative modeling. In text-to-spee

Bayesian Selective Latent Inference for Wastewater-First Influenza Monitoring

Model ReleasesDGX agent

arXiv:2606.09433v1 Announce Type: new Abstract: Wastewater influenza surveillance can reveal community circulation before clinical reporting, but wastewater alone is not a fully identifiable proxy for

BCG-FM: A Foundation Model for Ambient Cardiac Health Sensing

ResearchDGX agent

arXiv:2606.07692v1 Announce Type: cross Abstract: Foundation models for wearable biosignals have matched or exceeded supervised specialists across a range of clinical tasks, yet all rely on modalities


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

BEACON: Behavioral Entropy Aggregation for Cross-Model Hallucination Detection in Large Language Models

ResearchDGX agent

arXiv:2606.07528v1 Announce Type: cross Abstract: Hallucination in large language models (LLMs), defined as the generation of factually incorrect or unsupported content, remains a critical barrier to

Benchmarking Open-Ended Multi-Agent Coordination in Language Agents

Model ReleasesDGX agent

arXiv:2606.08340v1 Announce Type: new Abstract: As language models are increasingly deployed as autonomous agents, they must coordinate with others over long horizons in open-ended interactive tasks.

Benchmarking Vision-Language-Action Models on SO-101: Failure and Recovery Analysis

Model ReleasesDGX agent

arXiv:2606.08881v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated strong generalization in robotic manipulation, yet existing evaluations are primarily conducted

Beware of GeeksBearing Gifts: Building True EU Frontier AI Sovereignty

SafetyDGX agent

arXiv:2606.07536v1 Announce Type: cross Abstract: Frontier artificial intelligence is reshaping all aspects of society, from economic output or military capability to democratic institutions. The EU i

Beyond Accuracy: Interpreting Topic Representation in Suicide Ideation Detection Models

SafetyDGX agent

arXiv:2606.07714v1 Announce Type: cross Abstract: Suicide ideation detection models are typically evaluated using aggregate performance metrics, yet little is known about how they internally represent

Beyond Additivity: Causal Discovery in Location-Scale Noise Models with Hidden Variables

ResearchDGX agent

arXiv:2606.08196v1 Announce Type: cross Abstract: We study causal discovery from observational data when some variables are hidden and the data-generating process follows a location-scale noise model

Beyond Agent Architecture: Execution Assumptions and Reproducibility in LLM-Based Trading Systems

AgentsDGX agent

arXiv:2606.08285v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems are increasingly proposed for financial trading, yet their reported performance remains difficult to co

Beyond English benchmarks: clinical llm evaluation in Brazilian Portuguese

Model ReleasesDGX agent

arXiv:2606.07853v1 Announce Type: cross Abstract: Large Language Models are transforming the support for clinical decision and their application in real scenarios. Yet, most benchmarks are conducted i

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.07805v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) from passive assistants to autonomous, execution-capable agents has introduced critical operational

Beyond Humans: Multispecies Animal Face Recognition Using Transfer Learning

ResearchDGX agent

arXiv:2606.09353v1 Announce Type: cross Abstract: Individual animal recognition can be useful in the search for lost or stolen pets, the tracking of individuals of endangered species, and the recognit

Beyond Item IDs: Scaling Short-Form-Video Recommendation via Semantic-Native Long Sequence Modeling

ApplicationsDGX agent

arXiv:2606.07546v1 Announce Type: cross Abstract: Capturing user interests across extensive watch histories is critical for short-form video recommendation, yet scaling sequence length is limited by t

Beyond Pass Rate: A Multilingual, Execution-Grounded Evaluation of Open Code LLMs

Model ReleasesDGX agent

arXiv:2606.08840v1 Announce Type: new Abstract: Code generation models are typically compared using compact execution benchmarks and aggregate pass rates, but such summaries obscure how performance va

Beyond Pass/Fail: Using Process Mining to Understand How LLMs Resist (and Fail) Red Team Attacks

Model ReleasesDGX agent

arXiv:2606.07833v1 Announce Type: cross Abstract: Standard AI red teaming evaluations reduce adversarial campaigns to a single binary outcome, attack success rate (ASR), not taking into account the se

Beyond Point Estimates: Benchmarking Uncertainty Quantification Methods on the AION-1 Astronomical Foundation Model

Local AiDGX agent

arXiv:2606.07771v1 Announce Type: cross Abstract: Foundation models for astronomical surveys offer powerful learned representations that can be transferred to downstream regression tasks such as galax

Beyond Probabilistic Similarity: Structural, Temporal, and Causal Limitations of Retrieval-Augmented Generation in the Legal Domain

ApplicationsDGX agent

arXiv:2606.09724v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become a standard architectural response to unreliability in legal AI, yet high-profile failures, including fab

Bidirectional Semantic Complementary Tool Retrieval for Remote Sensing Agents

Model ReleasesDGX agent

arXiv:2606.07538v1 Announce Type: cross Abstract: Large language model (LLM)-based agents provide a novel paradigm for the automated processing of remote sensing(RS) data. Their success in complex RS

Bidirectional Small-Granularity Search between Code and Text

Model ReleasesDGX agent

arXiv:2606.07519v1 Announce Type: cross Abstract: We introduce the novel task of bidirectional small-granularity search between code and text, where the queries are small snippets of text or code and

BioVid: Autoregressive Video Generation with Biological Behavior Semantic Comprehension

Model ReleasesDGX agent

arXiv:2606.08674v1 Announce Type: cross Abstract: Existing video generation frameworks treat sequence duration as an externally prescribed parameter -- fixed frame counts or text prompts -- producing

BLM-SGAN: Bidirectional Language Modeling for Semantic-Spatial Text-to-Image Generation

ResearchDGX agent

arXiv:2606.08847v1 Announce Type: cross Abstract: Despite the success of image generation from text descriptions, it still faces challenges that are difficult to overcome in domains such as natural la

Blockchain Infrastructure for Intelligent Cyber--Physical--Social Systems:Post-Quantum Security, Interoperability, and Trustworthy Data Economies in the Era of Embodied AI

TutorialsDGX agent

arXiv:2606.06895v1 Announce Type: cross Abstract: The deployment of embodied artificial intelligence via world-model-based robotics presents a transformative opportunity for blockchain infrastructure,

BRAIN: Bayesian Reasoning via Active Inference for Agentic and Embodied Intelligence in Mobile Networks

HardwareDGX agent

arXiv:2602.14033v1 Announce Type: cross Abstract: Future sixth-generation (6G) mobile networks will demand artificial intelligence (AI) agents that are not only autonomous and efficient, but also capa

Brain-Prompt Injection: A Route-Safety Audit for BCI-LLM Agents

SafetyDGX agent

arXiv:2606.09315v1 Announce Type: cross Abstract: BCI-to-agent pipelines turn decoded neural activity into an authorization channel for tool-use agents, exposing a new attack surface we call brain-pro

Brain2Text Decoding Model Reveals the Neural Mechanisms of Visual Semantic Processing

ResearchDGX agent

arXiv:2503.22697v3 Announce Type: replace-cross Abstract: Decoding sensory experiences from neural activity to reconstruct human-perceived visual stimuli and semantic content remains a challenge in ne

Bridging Expert Knowledge and Automated Feature Engineering via Self-Evolution

SafetyDGX agent

arXiv:2606.08800v1 Announce Type: new Abstract: In high-stakes settings such as brand compliance, clinical care, and content moderation, machine learning cannot be deployed as opaque oracles: practiti

Bridging Traditional Explainability Methods and Multimodal Multilingual Models: An XAI-Based Analysis

SafetyDGX agent

arXiv:2606.07533v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) effectively integrate text and audio to interpret context in complex interactive dialogues. However, the inte

BSTabDiff: Block-Subunit Diffusion Priors for High-Dimensional Tabular Data Generation

Model ReleasesDGX agent

arXiv:2606.09257v1 Announce Type: cross Abstract: High-Dimensional Low-Sample Size (HDLSS) tabular domains (e.g., omics) are characterized by n ll m, where n = number of samples, and m = number of fea

Calibration of Structured Ignorance Certificates for Diagnosing Unknown Unknowns in Reasoning Models

Model ReleasesDGX agent

arXiv:2606.08571v1 Announce Type: cross Abstract: Large language models frequently fail in a characteristic way: rather than acknowledging ignorance, they produce fluent but incorrect answers to quest

Can Data Work be Reparative?

SafetyDGX agent

arXiv:2606.09408v1 Announce Type: cross Abstract: We present an ethnographic study of an alternative approach to data work, developed by a civic-tech initiative that builds datasets for training and b

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP

Model ReleasesDGX agent

arXiv:2505.11189v3 Announce Type: replace Abstract: Large language models (LLMs) can amplify misinformation, undermining societal goals such as the UN SDGs. We study three documented drivers of misinf

Can the Environment Speak for Itself? T^{2}-GRPO: A Turn-Trajectory Group Relative Policy Optimization for Caregiver Agents

SafetyDGX agent

arXiv:2606.08875v1 Announce Type: new Abstract: Optimizing large language models (LLMs) for long-horizon caregiver agents requires balancing delayed task objectives with immediate environment dynamics

Can You Trust What You See? Human and AI Detection of Synthetic Legal Evidence

Model ReleasesDGX agent

arXiv:2606.07613v1 Announce Type: cross Abstract: Visual evidence has long been treated as a reliable form of legal proof, but advances in artificial intelligence (AI) are undermining that assumption.

CANS: Accelerating Multiuser Collaborative Edge Inference via Cooperative Autodidactic NeuroSurgeon

Local AiDGX agent

arXiv:2606.09175v1 Announce Type: cross Abstract: Recently, mobile edge computing (MEC)-enabled collaborative deep neural network (DNN) inference has emerged as a promising approach for delivering int

Capability-Aligned Hierarchical Learning for Tool-Augmented LLMs

SafetyDGX agent

arXiv:2606.09371v1 Announce Type: new Abstract: Tool learning enables LLMs to invoke external tools to accomplish tasks. Prior studies have demonstrated the effectiveness of a hierarchical structure:

Capacity, Not Format: Rethinking Structured Reasoning Failures

ResearchDGX agent

arXiv:2606.09410v1 Announce Type: new Abstract: Prior work treats structured output as a reasoning tax, but this framing is incomplete: the cost of formatting depends strongly on a model's spare capac

CAPruner: Conceptual-Adjacent Scene Graph Pruner for Enhancing 3D Spatial Reasoning of Large Language Models

ResearchDGX agent

arXiv:2606.07529v1 Announce Type: cross Abstract: Large language models (LLMs) have recently been applied to 3D vision-language (3D-VL) tasks, which require spatial reasoning to identify target object

CARE: A Conformal Safety Layer for Medical Summarization

SafetyDGX agent

arXiv:2606.08969v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for medical summarization, but their outputs can omit medically important information and introduce

Causal Agent Replay: Counterfactual Attribution for LLM-Agent Failures

Model ReleasesDGX agent

arXiv:2606.08275v1 Announce Type: cross Abstract: When an LLM agent fails -- issues a refund it should not have, calls the wrong tool, leaks data -- existing tooling answers what happened (observabili

CausShield: Sample Reconstruction-Resilient Vertical FL via Causal Representation Learning

ResearchDGX agent

arXiv:2606.08027v1 Announce Type: cross Abstract: Vertical federated learning (VFL) is a distributed learning paradigm that leverages vertically partitioned features across isolated parties without sh

Cheap Reward Hacking Detection

ResearchDGX agent

arXiv:2606.08893v1 Announce Type: cross Abstract: A small transformer encoder is trained to map Terminal-Wrench trajectories onto a unit sphere where embedding distance approximates the L_1 distance b

Cherry-pick Override: Unsafe Directional Commitment in LLM Judges under Mixed Evidence

ResearchDGX agent

arXiv:2606.07834v1 Announce Type: cross Abstract: LLM judges increasingly turn verdicts into system commitments. Under mixed evidence (claims with both supporting and refuting sources) this is unsafe:

Chiaroscuro Attention: Spending Compute in the Dark

ResearchDGX agent

arXiv:2606.08327v1 Announce Type: cross Abstract: Standard transformers apply self-attention uniformly at every layer and token, regardless of whether the input requires dynamic cross-token interactio

CHIMERA-Bench: A Benchmark Dataset for Epitope-Specific Antibody Design

Model ReleasesDGX agent

arXiv:2603.13431v3 Announce Type: replace-cross Abstract: Computational antibody design has seen rapid methodological progress, with dozens of deep generative methods proposed in the past three years,

CLASP: Language-Driven Robot Skill Selection and Composition using Task-Parameterized Learning

Model ReleasesDGX agent

arXiv:2606.08169v1 Announce Type: cross Abstract: Enabling robots to understand and execute tasks from natural language commands while maintaining data efficiency remains challenging. Foundation model

CLONE: A 3DGS-Based Closed-Loop Differentiable Optimization Framework for Single-Image Normal Estimation

ResearchDGX agent

arXiv:2508.05950v2 Announce Type: replace-cross Abstract: We propose CLONE, a 3DGS-based Closed-Loop differentiable Optimization framework for single-image Normal Estimation. The core idea is to const

Closing the Prior-Posterior Loop: Self-Reflective Molecular Design with Analysis-Driven LLM Iteration

ResearchDGX agent

arXiv:2606.09520v1 Announce Type: cross Abstract: Can a general-purpose large language model design molecules with the precision of a seasoned chemist? Current LLM-based frameworks answer this questio

Closing the Sim-to-Real Gap: An Evaluation Framework for Autonomous Cyber Defense Configuration of Commercial EDR

Model ReleasesDGX agent

arXiv:2606.08168v1 Announce Type: cross Abstract: Leading commercial endpoint detection and response (EDR) products have shifted from operator-configured rule sets to multi-component systems where aut

Closure-Validated Circuit Discovery in Attention Heads: Co-activation Proposes, Ablation Disposes

ResearchDGX agent

arXiv:2606.09607v1 Announce Type: cross Abstract: Interpretability increasingly treats groups of components, not individual units, as the basic object, and proposes to find them by clustering co-activ

CLPO: Curriculum Learning meets Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2509.25004v2 Announce Type: replace Abstract: Online reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning abilities of large languag

CodeTaste: Can LLMs Generate Human-Level Code Refactorings?

Model ReleasesDGX agent

arXiv:2603.04177v2 Announce Type: replace-cross Abstract: LLM coding agents can generate working code, but their solutions often accumulate complexity, duplication, and architectural debt. Human devel

Collaborative Edge-to-Server Inference for Vision-Language Models

Local AiDGX agent

arXiv:2512.16349v2 Announce Type: replace-cross Abstract: We propose a collaborative edge-to-server inference framework for vision-language models (VLMs) that reduces communication cost while maintain

Collaborative Human-Agent Protocol (CHAP)

AgentsDGX agent

arXiv:2606.09751v1 Announce Type: new Abstract: Foundation models are moving from response generation into operational roles. They plan across steps, call tools, request human input, coordinate with o

Comparative evaluation of training strategies using partially labelled datasets for segmentation of white matter hyperintensities and stroke lesions in FLAIR MRI

SafetyDGX agent

arXiv:2601.20503v2 Announce Type: replace-cross Abstract: White matter hyperintensities (WMH) and ischaemic stroke lesions (ISL) are key imaging biomarkers of cerebral small vessel disease (SVD) detec

Complement or substitute? How AI increases the demand for human skills

ApplicationsDGX agent

arXiv:2412.19754v4 Announce Type: replace-cross Abstract: Artificial Intelligence (AI) is transforming the nature of work, yet there is limited empirical evidence on how it affects demand for human sk

ComplexConstraints and Beyond: Expert Rubrics for RLVR

Model ReleasesDGX agent

arXiv:2606.09118v1 Announce Type: new Abstract: As LLM capabilities advance rapidly, the evaluation methods used to assess them increasingly lag behind. Traditional benchmarks relied on programmatic v

Component Ablation for Efficient Hybrid Language Model Architectures: Performance, Resilience, and Compression Implications

Model ReleasesDGX agent

arXiv:2603.22473v2 Announce Type: replace-cross Abstract: Hybrid language models combine softmax attention with linear-time sequence mechanisms such as state-space or linear-attention layers, but the

Conan-embedding-v3: Fusing Modality-Specific Models for Omni-Modal Embedding

Model ReleasesDGX agent

arXiv:2606.09331v1 Announce Type: cross Abstract: Omni-modal retrieval promises a single embedding space for text, image, video, document, and audio inputs, but building such a unified retriever is di

Concerns and Strategic Responses of Older Workers Navigating Generative AI in Bridge Employment

ResearchDGX agent

arXiv:2606.07543v1 Announce Type: cross Abstract: Generative AI (GenAI) is transforming workplaces at a rapid pace. This disproportionately affects vulnerable communities, including older workers (OWs

← Previous
1…141142143144145…358
Next →