AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
7 Jul 2026

Attention Limited Reward Learning

SafetyDGX agent

arXiv:2607.04590v1 Announce Type: new Abstract: Pairwise human comparisons are a primary interface through which modern AI systems learn human preferences. RLHF and related alignment pipelines typical

Attributing Emergence in Million-Agent Systems

SafetyDGX agent

arXiv:2605.11404v2 Announce Type: replace Abstract: Large language models (LLMs) can simulate human-like reasoning and decision-making in individual agents. LLM-powered multi-agent systems (MAS) combi

Auto-AEG: Scalable Data Construction for Open-Vocabulary Audio Event Grounding

Model ReleasesDGX agent

arXiv:2607.04383v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) reason fluently about sound yet struggle to localize precisely when events occur, while classical Sound Event Dete


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Auto: The AGI Compiler

Model ReleasesDGX agent

arXiv:2607.04542v1 Announce Type: cross Abstract: Every LLM agent run re-derives its behavior token by token on a frontier model: brilliant, expensive, slow, and unbounded. We present Auto, a compiler

AutoCedar: An Agentic Framework for Verifier-Guided Access Control Policy Synthesis

Model ReleasesDGX agent

arXiv:2607.03656v1 Announce Type: cross Abstract: Large Language Models are increasingly used to turn natural-language requirements into code. In access control, that shortcut is dangerous: a generate

Automated Data Readiness for Scientific AI

AgentsDGX agent

arXiv:2607.02771v1 Announce Type: new Abstract: Leadership computing facilities steward large-scale scientific datasets that routinely require substantial transformation before serving as AI training

AutoResearch: An Execution-Grounded Multi-Agent Framework for Reliable Research Workflow Automation

Model ReleasesDGX agent

arXiv:2607.02520v1 Announce Type: cross Abstract: Automated research agents increasingly generate code, retrieve literature, and draft scientific artifacts, but they often fail to verify whether gener

AViS-Mamba: Adaptive Visual Steering of Audio State-Space Dynamics for Violence Detection

SafetyDGX agent

arXiv:2604.03329v2 Announce Type: replace-cross Abstract: Automatic violence detection from video is challenging because violent interactions may be distant, occluded, or only partially visible. Audio

Back to Basics: Improving Molecular Understanding in LLMs via SMILES-Graph Translation

Model ReleasesDGX agent

arXiv:2607.03007v1 Announce Type: cross Abstract: Recent advances in molecular large language models have led to strong performance on molecular understanding and generation tasks, yet these gains oft

BanglaMemeEvidence: A Multimodal Benchmark Dataset for Explanatory Evidence Detection in Bengali Memes

Model ReleasesDGX agent

arXiv:2607.03981v1 Announce Type: cross Abstract: Memes have become influential communication tools on social media, combining viral visuals with concise messaging to convey impactful ideas. While sub

Benchmarking API Drift in LLM-Generated Quantum Code Across Successive SDK Versions

Model ReleasesDGX agent

arXiv:2607.04072v1 Announce Type: cross Abstract: Large language models can generate plausible quantum code, but it is unclear whether they can reliably target the specific software development kit (S

Best-of-Better-N: Generating Pre-Aligned Responses with In-Context Learning

SafetyDGX agent

arXiv:2607.03453v1 Announce Type: cross Abstract: Inference-time alignment methods, such as Best-of-N, offer a flexible alternative to training-based alignment by using reward models to select high-qu

BEVLM: Distilling Semantic Knowledge from LLMs into Bird's-Eye View Representations

SafetyDGX agent

arXiv:2603.06576v2 Announce Type: replace-cross Abstract: The integration of Large Language Models (LLMs) into autonomous driving has attracted growing interest for their strong reasoning and semantic

Beyond Forecasting: The Belief-to-Trade Layer in Prediction-Market Agents

Model ReleasesDGX agent

arXiv:2607.03015v1 Announce Type: new Abstract: Forecasting future events has attracted growing attention as a testbed for general-purpose AI. A natural way to ground this evaluation is let the models

Beyond Independent Labels: Schwartz-Geometry Decoding for Human Value Detection

SafetyDGX agent

arXiv:2607.05052v1 Announce Type: cross Abstract: Human value detection is commonly formulated as sentence-level multi-label classification over the 19 refined Schwartz values, typically predicted as

Beyond Multilingual Averages: MTEB-PT, a Benchmark for Portuguese Sentence Encoders

Model ReleasesDGX agent

arXiv:2607.04071v1 Announce Type: cross Abstract: Portuguese remains underrepresented in text embedding evaluation, despite being one of the most widely spoken languages in the world. As a result, emb

Beyond Static Rules: Automated Discovery of Latent Vulnerabilities in Text-to-SQL

ApplicationsDGX agent

arXiv:2607.03833v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved remarkable success in Text-to-SQL tasks, their deployment in real-world environments is hindered by l

Beyond Task Completion: A Verification-vs.-Conformance Gap in Tool-Evolving Agents

Model ReleasesDGX agent

arXiv:2604.00392v2 Announce Type: replace-cross Abstract: Agents that synthesize their own tools ship a second artifact alongside each answer: a software library that future tasks reuse, compose, and

Biological Motifs for Agentic Control

AgentsDGX agent

arXiv:2607.04240v1 Announce Type: new Abstract: The transition of Large Language Models (LLMs) from passive generators to autonomous agents has introduced significant challenges in reliability, securi

Boosting Automatic Exercise Evaluation Through Musculoskeletal Simulation-Based IMU Data Augmentation

ApplicationsDGX agent

arXiv:2505.24415v2 Announce Type: replace-cross Abstract: Automated evaluation of movement quality can enhance physiotherapeutic treatment and sports training by providing objective, real-time feedbac

Bootstrap Flow-Map Tree Sampling Enables Online Feedback Driven Search

SafetyDGX agent

arXiv:2607.02915v1 Announce Type: cross Abstract: In many scientific and engineering domains, maximizing discovery within a limited sampling budget demands strategic, observation-guided exploration. W

BoRP: Bootstrapped Regression Probing for Scalable and Human-Aligned LLM Evaluation

SafetyDGX agent

arXiv:2601.18253v2 Announce Type: replace-cross Abstract: Accurate evaluation of user satisfaction is critical for iterative development of conversational AI. However, for open-ended assistants, tradi

Brand-as-Memory: Vision-Language Models Encode Causal, Mechanistically Localizable Credibility Priors for News Sources

Model ReleasesDGX agent

arXiv:2607.03365v1 Announce Type: cross Abstract: Vision-language models (VLMs) increasingly read news and web content as images, where the publisher's identity is visually present. We show that VLMs

Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process

SafetyDGX agent

arXiv:2607.03748v1 Announce Type: new Abstract: Unified multi-modal models (UMMs) have shown promising interleaved text-image reasoning capabilities, yet effectively optimizing such multi-turn generat

Builder, Defender, Breaker: The Case Against Removing the Human from the AI-Driven Security Lifecycle

AgentsDGX agent

arXiv:2607.03215v1 Announce Type: cross Abstract: Artificial intelligence has spread across the whole of the security lifecycle. The same family of models now writes application code, hardens it, and

CABTO: Context-Aware Behavior Tree Grounding for Robot Manipulation

ResearchDGX agent

arXiv:2603.16809v2 Announce Type: replace-cross Abstract: Behavior Trees (BTs) offer a powerful paradigm for designing modular and reactive robot controllers. BT planning, an emerging field, provides

CAGE-1: Control, Assurance, and Governance Evaluation for Enterprise Agentic AI

SafetyDGX agent

arXiv:2607.03510v1 Announce Type: cross Abstract: Enterprise artificial intelligence is moving from experimentation into operational workflows. Early programs focused on model access and retrieval-aug

Can Conversational Temporal Dynamics Improve Depression Detection in Dyads? A Preliminary Investigation in Multi-Modality Perspectives

ResearchDGX agent

arXiv:2607.03744v1 Announce Type: new Abstract: Automatic depression detection from clinical interviews typically models the semantic content and acoustic characteristics of participant speech. Howeve

Can Model Merging Improve Aggregation in DiLoCo?

Local AiDGX agent

arXiv:2607.03011v1 Announce Type: cross Abstract: Model merging techniques, which aggregate independently finetuned models into one to combine their capabilities, have become a topic of significant in

CanniUplift: A Holistic Framework for Mitigating Seller and Incentive Cannibalization in E-commerce Uplift Modeling

SafetyDGX agent

arXiv:2607.05242v1 Announce Type: cross Abstract: Personalized incentive allocation is vital for e-commerce, where uplift modeling is the standard for estimating Individual Treatment Effects (ITE). Ho

CaresAI at SMM4H-HeaRD 2026: Predicting TNM Staging

ApplicationsDGX agent

arXiv:2607.03466v1 Announce Type: cross Abstract: This study aims to predict Tumor, Node, and Metastasis (TNM) stage labels independently, with the Cancer Genome Atlas (TCGA) pathology report as the s

CARL: Constraint-Aware Reinforcement Learning for Planning with LLMs

ApplicationsDGX agent

arXiv:2607.04854v1 Announce Type: new Abstract: Despite their strong reasoning capabilities and extensive world knowledge, Large Language Models (LLMs) frequently generate plans that violate task cons

Causal Mechanism Reduction: Mechanism Replacement for Neural Network Pruning and Abstraction

SafetyDGX agent

arXiv:2602.24266v2 Announce Type: replace-cross Abstract: Which internal mechanisms of a neural network can be replaced while preserving the computation it performs? Structured pruning asks for smalle

CausalChaos! Dataset for Comprehensive Causal Action Question Answering Over Longer Causal Chains Grounded in Dynamic Visual Scenes

ResearchDGX agent

arXiv:2404.01299v3 Announce Type: replace-cross Abstract: Causal video question answering (QA) has garnered increasing interest, yet existing datasets often lack depth in causal reasoning. To address

CausalGame: Benchmarking Causal Thinking of LLM Agents in Games

Model ReleasesDGX agent

arXiv:2607.04293v1 Announce Type: cross Abstract: Building AI Scientist agents with Large Language Models (LLMs) has recently attracted growing attention. Since scientific discovery fundamentally reli

CGGS: Consistency-Augmented Geometric Gaussian Splatting for Ego-centric 3D Scene Generation

ResearchDGX agent

arXiv:2607.03819v1 Announce Type: cross Abstract: Challenges remain in ego-centric 3D scene generation due to limited view overlap and the dominant influence of individual perspectives on scene interp

ChainReaction: Causal Chain-Guided Reasoning for Modular and Explainable Causal-Why Video Question Answering

ResearchDGX agent

arXiv:2508.21010v3 Announce Type: replace-cross Abstract: Existing Causal-Why Video Question Answering (VideoQA) models often struggle with higher-order reasoning, relying on opaque, monolithic pipeli

Chronos: The AI Co-Historian

Model ReleasesDGX agent

arXiv:2604.03553v2 Announce Type: replace Abstract: AI is increasingly supporting, accelerating, and automating scientific discovery across subjects. Yet, the adoption of AI in historical research rem

CineMobile: On-Device Image-to-Video Diffusion for Cinematic Camera Motion Generation

Model ReleasesDGX agent

arXiv:2607.03803v1 Announce Type: cross Abstract: The growing demand for image-to-video creation on mobile devices has increasingly focused on cinematic motion effects like bullet time, dolly zoom, sl

ClassicLogic: A Knowledge-Driven Benchmark of Classic Puzzle Games for Evaluating Compositional Generalization

Model ReleasesDGX agent

arXiv:2607.05185v1 Announce Type: new Abstract: Compositional generalization, the ability to understand and produce novel combinations of known components, remains a fundamental challenge for modern a

ClinOCR-Bench: A Comprehensive Clinical Scanned Document Dataset for Optical Character Recognition Model Evaluation

Model ReleasesDGX agent

arXiv:2607.03650v1 Announce Type: cross Abstract: Extracting textual information from scanned medical documents, such as external laboratory reports and manually filled forms, has been a major challen

CoACT: Action-Preserving Observation Compression for Coding Agents

AgentsDGX agent

arXiv:2607.02911v1 Announce Type: cross Abstract: LLM-based coding agents solve software-engineering tasks through iterative interactions with development environments, where returned observations acc

Code Benchmarks Should Prioritize Rigor, Reliability, and Reproducibility

Model ReleasesDGX agent

arXiv:2501.10711v5 Announce Type: replace-cross Abstract: Code-related benchmarks play a critical role in evaluating large language models (LLMs), yet their quality fundamentally shapes how the commun

CoGen3D: An Agentic Human-AI Co-Design Pipeline for 3D Asset Generation for Virtual Reality

AgentsDGX agent

arXiv:2607.03731v1 Announce Type: cross Abstract: Creating 3D assets for virtual reality requires modeling expertise, which restricts the authorship of immersive experiences. Existing generative AI to

COMET: Combinatorial Optimization for Multiplex Editing Targets Via Constraint-Preserving QAOA

TutorialsDGX agent

arXiv:2607.02622v1 Announce Type: cross Abstract: Multiplex CRISPR-Cas9 gene editing requires selecting one guide RNA per target gene subject to cross-gene interactions: a constrained combinatorial pr

Comparison of Loss Functions for Robust Deep Learning-based Echocardiography Segmentation when Learning with Partially Labelled Data from Multiple Domains

ResearchDGX agent

arXiv:2607.05008v1 Announce Type: cross Abstract: Echocardiography is the first imaging modality used for assessing cardiac function, and accurate segmentation of cardiac structures is essential for d

Compressing the Validation Bottleneck: An Agentic Self-Driving Lab for Scientific Discovery

AgentsDGX agent

arXiv:2607.04508v1 Announce Type: new Abstract: Agentic AI-for-Science can automate ideation, planning, and analysis, but final validation still depends on real experiments. A self-driving lab (SDL) c

Conditional Clifford-Steerable CNNs for PDE Modeling

ResearchDGX agent

arXiv:2510.14007v2 Announce Type: replace-cross Abstract: We introduce Conditional Clifford-Steerable CNNs (C-CSCNNs), a unified framework that incorporates equivariance to arbitrary pseudo-Euclidean

Conditional Diffusion Guided Knowledge Transfer for Multi-Domain Knowledge Graph Completion

ResearchDGX agent

arXiv:2607.03154v1 Announce Type: cross Abstract: Multi-domain knowledge graph completion (MKGC) aims to improve missing triple prediction in a target KG by transferring knowledge from other support K

Conflict-Based Lazy Search for Fast Multi-Manipulator Planning

AgentsDGX agent

arXiv:2607.04124v1 Announce Type: cross Abstract: Employing multiple manipulators can boost efficiency and accomplish tasks that a single manipulator cannot do. However, real-time planning for multipl

CONFLUX: A Latent Diusion Model for 3D Chest-CT Synthesis with RL Post-Training

SafetyDGX agent

arXiv:2607.02998v1 Announce Type: cross Abstract: Controllable generative models of 3D medical images can synthesize volumes with specified clinical attributes, but this demands samples that are simul

Consistent but Miscalibrated: Evaluating LLM Limitations for Risk Communication in Natural Language

ResearchDGX agent

arXiv:2607.03882v1 Announce Type: cross Abstract: LLMs are increasingly deployed as post-hoc explainers of AI-generated outputs, yet it remains unclear whether they can reliably communicate probabilis

Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs

SafetyDGX agent

arXiv:2508.10031v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) have shown significant advancements in performance, various jailbreak attacks have posed growing safety and

Context Tuning for In-Context Optimization

ResearchDGX agent

arXiv:2507.04221v3 Announce Type: replace-cross Abstract: We introduce Context Tuning, a simple and effective method to significantly enhance few-shot adaptation of large language models (LLMs) withou

CONTRA: Red-Teaming Configurations of Personalizable Agents

SafetyDGX agent

arXiv:2607.03220v1 Announce Type: cross Abstract: Recent tools such as OpenClaw have extended the capabilities of LLM-based agents from simple dialog-based systems to fully autonomous agents. These sy

Correlation-Weighted Multi-Reward Optimization for Compositional Generation

ResearchDGX agent

arXiv:2603.18528v2 Announce Type: replace Abstract: Text-to-image models produce images that align well with natural language prompts, but compositional generation has long been a central challenge. M

Cortex: A Bidirectionally Aligned Embodied Agent Framework for Long-horizon Manipulation

AgentsDGX agent

arXiv:2607.05377v1 Announce Type: cross Abstract: While recent Vision-Language-Action (VLA) models show promise toward generalist manipulation policies, they struggle with long-horizon tasks due to th

Covert Trait Propagation Is Representation Alignment: Mechanistic Evidence from Hidden-Channel Distillation

SafetyDGX agent

arXiv:2607.04432v1 Announce Type: cross Abstract: A student model trained on pure uniform noise can still inherit its teacher's digit-classification ability, provided the two share initialization. Pre

CP-WSP: A Declarative CP-SAT Framework for Configurable Multi-Constraint Workforce Scheduling

Model ReleasesDGX agent

arXiv:2607.05177v1 Announce Type: new Abstract: Workforce scheduling is an NP-hard combinatorial optimization problem requiring simultaneous satisfaction of labor regulations, coverage requirements, e

CRISP: A Spatiotemporal Camera-Radar Backbone for Driving via Forecasting-Based World-Model Pretraining

AgentsDGX agent

arXiv:2607.04541v1 Announce Type: cross Abstract: Camera-radar (CR) fusion is a practical sensing configuration for autonomous driving, but existing models are typically trained with task-specific sup

← Previous
1…8485868788…358
Next →