AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
1 Jul 2026

BEST-RQ-2: Contextualize-Then-Predict, a Two-Step Approach for Self-Supervised Audio Representations

ResearchDGX agent

arXiv:2606.30700v1 Announce Type: cross Abstract: Self-supervised learning enables audio representations that transfer across domains and tasks. We present BEST-RQ-2, an evolution of BEST-RQ that reta

Better Understanding, Understanding Better

AgentsDGX agent

arXiv:2606.31892v1 Announce Type: cross Abstract: 'Any fool can know; the point is to understand.' A well-known remark often attributed to Einstein captures a widely shared intuition: understanding is

Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models

Model ReleasesDGX agent

arXiv:2606.31338v1 Announce Type: cross Abstract: Recent music audio-language models achieve high accuracy on instrument question-answering benchmarks, but it remains unclear whether this reflects rob


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Beyond But-for Test: Counterfactual Explanation in Abstract Argumentation via Actual Causality (Extended Version)

ResearchDGX agent

arXiv:2606.31080v1 Announce Type: cross Abstract: Counterfactual explanation in abstract argumentation calls for an answer to the what-if query: would the topic argument still be accepted if the statu

Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization

Model ReleasesDGX agent

arXiv:2606.31002v1 Announce Type: new Abstract: Theorem-proving benchmarks evaluate proof search against fixed formal statements, but natural-language-to-Lean formalization must generate the formal st

Beyond expert users: agents should help users construct preferences, not just elicit them

Model ReleasesDGX agent

arXiv:2606.30863v1 Announce Type: new Abstract: Agents typically assume an expert user -- one with well-formed preferences about what they want -- and default to clarifying questions whenever the task

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics

AgentsDGX agent

arXiv:2606.31134v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated exceptional capabilities in mathematical reasoning, they frequently produce subtle errors that evad

BP-TTA: Balanced and Prototype-Guided Test-Time Adaptation in Dynamic Scenarios

SafetyDGX agent

arXiv:2606.31420v1 Announce Type: new Abstract: Test-Time Adaptation (TTA) enables models trained on a source domain to adapt online to unlabeled test data under distribution shifts. While recent TTA

Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning

SafetyDGX agent

arXiv:2606.31825v1 Announce Type: cross Abstract: Recent multimodal large language models have shown great promise in clinical image reasoning, but existing post-training pipelines remain predominantl

Bridging Local Observation and Global Simulation in Closed-Loop Traffic Modeling

Local AiDGX agent

arXiv:2606.31844v1 Announce Type: cross Abstract: A local-to-global context mismatch arises when autoregressive traffic simulators trained on ego-centric driving logs are deployed in globally observab

Budget-Adaptive Routing: Skipping the Weak When the Strong Answers Anyway

ResearchDGX agent

arXiv:2606.30919v1 Announce Type: cross Abstract: Edge-cloud inference collaborations are often designed with a routing estimator that decides whether to offload each frame from weak models at the edg

Calibrating the Evaluator: Does Probability Calibration Mitigate Preference Coupling in LLM Agent Feedback Loops?

Model ReleasesDGX agent

arXiv:2606.31371v1 Announce Type: cross Abstract: When large language model (LLM) agents adapt their behavior through evaluator feedback, systematic evaluator biases propagate into the agent's learned

Can LLMs Imagine Moral Alternatives Beyond Binary Dilemmas?

AgentsDGX agent

arXiv:2606.31213v1 Announce Type: cross Abstract: As large language models (LLMs) are increasingly deployed as moral advisors and agents, they need to address dilemmas between two competing values. Ho

Can Physician Expertise Improve Machine Learning Identification of Delirium?

Model ReleasesDGX agent

arXiv:2606.30651v1 Announce Type: cross Abstract: Delirium is common in hospitalized patients and is often missed in routine care. We present a user-centered interactive machine learning (UC-iML) fram

Can VLMs Reason Robustly? A Neuro-Symbolic Investigation

ResearchDGX agent

arXiv:2603.23867v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have been applied to a wide range of reasoning tasks, yet it remains unclear whether they can reason robustly un

CDR-Bench: Evaluating Faithful Execution of Compositional, Order-Sensitive Data Refinement Recipes

Model ReleasesDGX agent

arXiv:2606.31435v1 Announce Type: new Abstract: Data refinement involves executing multi-step recipes over evolving text states, where both composition and execution order of processing operators dete

CharDiff-LP: A Diffusion Model with Character-Level Guidance for License Plate Image Restoration

ResearchDGX agent

arXiv:2510.17330v3 Announce Type: replace-cross Abstract: License plate image restoration is important not only as a preprocessing step for license plate recognition but also for enhancing evidential

CHERRY: Compressed Hierarchical Experts with Recurrent Representational Yield

Model ReleasesDGX agent

arXiv:2606.31796v1 Announce Type: cross Abstract: We study three complementary techniques for training compute-efficient language models. (1) Selective supervision and per-token efficiency. Selective

Citation Discipline in Spec-Driven Development: A Cross-Model Empirical Study of Output Determinism and Automated Hallucination Detection in LLM-Generated Code

Model ReleasesDGX agent

arXiv:2606.30689v1 Announce Type: cross Abstract: Spec-Driven Development (SDD) frameworks guide Large Language Model (LLM)-powered code generation through formal specifications, yet they differ funda

ClawArena-Team: Benchmarking Subagent Orchestration and Dynamic Workflows in Language-Model Agents

Model ReleasesDGX agent

arXiv:2606.31174v1 Announce Type: new Abstract: Production large language-model (LLM) agents are increasingly deployed not as lone problem-solvers but as managers: a main model creates specialized sub

CLIMB: Centroid-Based Hierarchical Memory for Online Continual Self-Supervised Learning

TutorialsDGX agent

arXiv:2606.31275v1 Announce Type: cross Abstract: Online Continual Self-Supervised Learning (OCSSL) aims to learn representations from a continuous stream of unlabeled data, without knowledge of task

CLOUDADV: Decision-Aligned Instance Sizing with Zero-Shot Foundation Models under Drift

SafetyDGX agent

arXiv:2606.31470v1 Announce Type: new Abstract: Cloud virtual machines are often overprovisioned, creating avoidable cost and operational inefficiency. We present CLOUDADV, an interactive engineer-fac

ComAct: Reframing Professional Software Manipulation via COM-as-Action Paradigm

Model ReleasesDGX agent

arXiv:2606.13239v2 Announce Type: replace-cross Abstract: Existing computer-use agents remain fundamentally limited in professional software manipulation: GUI-based agents suffer from fragile visual g

Comparative Analysis of Machine Learning based Intrusion Detection in Realistic IoT Networks

ApplicationsDGX agent

arXiv:2606.31594v1 Announce Type: cross Abstract: The Internet of Things (IoT) is rapidly growing and expanding into various sectors, such as healthcare, transportation, smart homes, and more. Despite

ComplianceGate: Classifier-Gated Multi-Tier LLM Routing for Inference in Regulated Industries

Local AiDGX agent

arXiv:2606.31163v1 Announce Type: cross Abstract: Large language models deployed in regulated industries operate under two constraints: compliance enforcement and cost efficiency. Personally identifia

Compositional Concept-Based Neuron-Level Interpretability for Deep Reinforcement Learning

AgentsDGX agent

arXiv:2502.00684v2 Announce Type: replace-cross Abstract: Deep reinforcement learning (DRL) has successfully addressed many complex control problems. However, the neural networks representing policies

Contrastive Reflection for Iterative Prompt Optimization

AgentsDGX agent

arXiv:2606.30840v1 Announce Type: new Abstract: LLM agents are becoming central to information retrieval: they issue retrieval queries, synthesize answers, and increasingly serve as judges for IR eval

CoReLIN: Constraint-based Reasoning for Zero-shot Lifelong Interactive Navigation

ApplicationsDGX agent

arXiv:2602.20055v2 Announce Type: replace-cross Abstract: Robot navigation typically assumes an obstacle-free path exists between start and goal. In real environments, however, clutter may block all r

Corruption Robust Offline Reinforcement Learning with Human Feedback

SafetyDGX agent

arXiv:2402.06734v2 Announce Type: replace-cross Abstract: We study data corruption robustness for reinforcement learning with human feedback (RLHF) in an offline setting. Given an offline dataset of p

Creating Intelligence: A Computational Foundation for AGI

ResearchDGX agent

arXiv:2606.31819v1 Announce Type: new Abstract: This work introduces a new computational theory of mind grounded in set theory and hyperdimensional computing. Whereas traditional neural networks rely

Cross-Domain Feature Expansion for Tabular Medical Data via Knowledge Graphs Injection

ResearchDGX agent

arXiv:2606.31171v1 Announce Type: new Abstract: Acquiring comprehensive cross-domain biomedical profiles is often costly and time-consuming, resulting in severe data scarcity in medical research. To a

Cross-lingual Relation Extraction with Large Language Models: Zero-Shot, Few-Shot, and Fine-Tuned Evaluation on Romanian

Model ReleasesDGX agent

arXiv:2606.31718v1 Announce Type: cross Abstract: Relation extraction (RE) for low-resource languages is typically constrained by the lack of annotated corpora. We investigate the feasibility of cross

Cross-Modal Hierarchical Fusion for from Multi-Sensor Ground Observation

ResearchDGX agent

arXiv:2606.30647v1 Announce Type: cross Abstract: Dense volumetric reconstruction of cloud microphysical fields from sparse ground-based instruments remains an open problem, largely because the availa

CryoACE: An Atom-centric Framework for Accurate and Automated Model Building in Cryo-EM

ApplicationsDGX agent

arXiv:2606.31332v1 Announce Type: new Abstract: Protein automodeling from cryo-EM density maps faces unique challenges in enforcing physicochemical validity and managing conformational heterogeneity.

CSO-LLM: Class Subspace Orthogonalization for Post-Training Backdoor Detection and Trigger Inversion in LLMs

ResearchDGX agent

arXiv:2606.31309v1 Announce Type: cross Abstract: While post-training backdoor detection and trigger inversion schemes have been developed for AIs used e.g. for images, there is a paucity of such meth

CSTrader: A Testbed for Language-Grounded Trading in a Community-Driven Virtual Asset Market

Model ReleasesDGX agent

arXiv:2606.31461v1 Announce Type: new Abstract: Niche asset markets, such as Counter-Strike 2 (CS2) weapon skins, are small, volatile, and heavily driven by community discussions and platform rules. T

Curvature-Guided Module Localization for Low-Rank Detoxification of Backdoored Large Language Models

Model ReleasesDGX agent

arXiv:2606.30899v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to large language models (LLMs) by causing otherwise benign systems to produce attacker-specified malicious beh

CVE-TTP KG: Knowledge Graph Linking Software Vulnerabilities to Attack Behaviors

ResearchDGX agent

arXiv:2606.31557v1 Announce Type: cross Abstract: In the evolving threat landscape, adversaries exploit software vulnerabilities to launch sophisticated attacks, challenging traditional defenses. Alth

DA-Studio: An Agentic System for End-to-End Data Analysis

AgentsDGX agent

arXiv:2606.31423v1 Announce Type: cross Abstract: Real-world data analysis is a multi-step process over heterogeneous inputs rather than merely producing a final answer. A practical system should auto

Dataset Construction for Training LLM to Learn Analog Circuit Knowledge

Model ReleasesDGX agent

arXiv:2508.10409v3 Announce Type: replace-cross Abstract: This paper constructs a textual dataset for training large language models (LLMs) to learn analog circuit knowledge and customizes LLM trainin

DDIAgents: Mechanism-Conditioned Context Flow for Drug-Drug Interaction Prediction

SafetyDGX agent

arXiv:2606.31085v1 Announce Type: new Abstract: Drug-drug interaction (DDI) prediction is essential for medication safety, yet it requires reasoning over heterogeneous biomedical evidence whose releva

Deductive Logic in Language Models: Horizontal vs Vertical Reasoning

TutorialsDGX agent

arXiv:2510.09340v2 Announce Type: replace Abstract: Recent language models exhibit significant logical reasoning abilities, yet the mechanisms supporting deductive inference remain poorly understood.

Delta-JEPA: Learning Action-Sensitive World Models via Latent Difference Decoding

ResearchDGX agent

arXiv:2606.31232v1 Announce Type: new Abstract: Learning visual world models for planning requires compact latent dynamics that remain sensitive to actions, yet reconstruction-free joint-embedding obj

Design and Implementation of Agentic Orchestrations and Orchestration of Agents

AgentsDGX agent

arXiv:2606.31518v1 Announce Type: new Abstract: Agentic Business Process Management has gained momentum recently. The prospect is that the autonomy of AI agents, i.e., predominantly LLM-based agents,

Detecting Audio Deepfakes on the Edge:Lightweight SSL-Based Detection in a Browser Plugin

Local AiDGX agent

arXiv:2606.30780v1 Announce Type: cross Abstract: Audio deepfakes are a growing challenge for the general public, as well as for journalists and fact-checkers. The latter need reliable tools to verify

DeXposure-Claw: An Agentic System for DeFi Risk Supervision

AgentsDGX agent

arXiv:2606.19501v2 Announce Type: replace Abstract: Decentralized finance exposes supervisors to fast-moving, networked credit risks. General-purpose LLM agents fit this setting poorly: they over-read

DeXposure-FM: A Time-series, Graph Foundation Model for Credit Exposures and Stability on Decentralized Financial Networks

ApplicationsDGX agent

arXiv:2602.03981v2 Announce Type: replace-cross Abstract: Credit exposure in Decentralized Finance (DeFi) is often implicit and token-mediated, creating a dense web of inter-protocol dependencies. Thu

Diffusion Crossover: Defining Evolutionary Recombination in Diffusion Models via Noise Sequence Interpolation

ResearchDGX agent

arXiv:2604.14790v2 Announce Type: replace Abstract: Interactive Evolutionary Computation (IEC) provides a powerful framework for optimizing subjective criteria such as human preferences and aesthetics

Disentangling Reasoning Logic to Resolve Explicit Knowledge Conflicts

Model ReleasesDGX agent

arXiv:2508.01273v3 Announce Type: replace Abstract: Explicit knowledge conflicts, occurring when retrieved contexts contain contradictory information, pose a fundamental challenge for Large Language M

Distilling Temporal Coherence into 2D Networks for Transrectal Ultrasound Prostate Video Segmentation

Model ReleasesDGX agent

arXiv:2606.31198v1 Announce Type: cross Abstract: Real-time video segmentation of the prostate in Transrectal Ultrasound (TRUS) is essential for image-guided interventions. While conventional 2D metho

Distilling the Essence: Efficient Reasoning Distillation via Sequence Truncation

ResearchDGX agent

arXiv:2512.21002v3 Announce Type: replace-cross Abstract: Distilling the capabilities from a large reasoning model (LRM) to a smaller student model often involves training on substantial amounts of re

DPPE: Rethinking Camera-Based Positional Encoding for Scaling Multi-View Transformers

ResearchDGX agent

arXiv:2606.31585v1 Announce Type: cross Abstract: The remarkable scalability of Transformers has expanded their application to 3D computer vision, where camera-aware positional encoding is crucial for

DSIP: A Dynamic Coordination Planner for Signal-Free Intersections using Diffusion-Model-Based Multi-Agent Motion Planning

AgentsDGX agent

arXiv:2606.30694v1 Announce Type: cross Abstract: Traffic signal control at urban intersections inherently introduces stop-and-go behavior, resulting in increased delays and reduced traffic efficiency

ECHO: Prune to act, trace to learn with selective turn memory in agentic RL

SafetyDGX agent

arXiv:2606.31650v1 Announce Type: cross Abstract: Long-horizon language agents must repeatedly interact with tools, accumulate evidence, and make decisions under bounded context windows. Existing cont

ELEVATE: Designing Human-Centered GenAI Virtual Tutors for Scalable and Inclusive Education

Local AiDGX agent

arXiv:2606.30662v1 Announce Type: cross Abstract: The advent of Generative Artificial Intelligence (GenAI), and in particular Large Language Models (LLMs), is reshaping educational practice, while int

Embodied CAD: Solver-Grounded LLM Agents for Parametric B-Rep Assembly Modeling

Model ReleasesDGX agent

arXiv:2606.31252v1 Announce Type: new Abstract: Large language models can write plausible CAD scripts, but reliable industrial CAD modeling requires more than syntactically valid code: every feature,

Emergent Culture in Minimal LLM Systems

AgentsDGX agent

arXiv:2606.30668v1 Announce Type: cross Abstract: What happens when LLM agents operate with no context outside a turn, minimal prompting, and simple tools? Inspired by swarm engineering, we give colle

Enhancing Graph Representations with Neighborhood-Contextualized Message-Passing

Model ReleasesDGX agent

arXiv:2511.11046v3 Announce Type: replace-cross Abstract: Graph neural networks (GNNs) have become an indispensable tool for analyzing relational data. Classical GNNs are broadly classified into three

Estimating the Effect of Timing on Coupon Effectiveness

ResearchDGX agent

arXiv:2606.30664v1 Announce Type: cross Abstract: The coupon incentive is one of the most common tools marketers use to court users to engage with a business at various stages of the customer life cyc

Evil Spectra: How Optimisers can Amplify or Suppress Emergent Misalignment

SafetyDGX agent

arXiv:2606.31591v1 Announce Type: cross Abstract: Emergent misalignment (EM) is a recently discovered phenomenon in LLMs where fine-tuning on a narrow misaligned task, such as writing insecure code, l

← Previous
1…103104105106107…358
Next →