AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
9 Jun 2026

Condition-Gated Reasoning for Context-Dependent Biomedical Question Answering

Model ReleasesDGX agent

arXiv:2602.17911v3 Announce Type: replace-cross Abstract: Current biomedical question answering (QA) systems often assume that medical knowledge applies uniformly, yet real-world clinical reasoning is

ConMem: Structured Memory-Guided Adaptation in Training-Free Multi-Agent Systems

AgentsDGX agent

arXiv:2606.08702v1 Announce Type: new Abstract: Recent advances have improved the adaptive capabilities of LLM-based multi-agent systems (MAS) through memory-, skill-, and learning-based approaches, y

Considerations for an Integrated Detector Design at FCC-ee: A Human-AI Exploration

ResearchDGX agent

arXiv:2606.07564v1 Announce Type: cross Abstract: This report explores detector design considerations for the Future Circular Collider in its electron-positron mode (FCC-ee) through an extended dialog


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Constrained Paraphrase Consistency for LLM Hallucination Detection

SafetyDGX agent

arXiv:2606.08158v1 Announce Type: cross Abstract: Large language models (LLMs) can generate factually inconsistent claims, motivating accurate and scalable hallucination detectors. Prior work largely

Contemporary AI lacks the imagination to diverge or negate in science

ResearchDGX agent

arXiv:2606.08251v1 Announce Type: cross Abstract: Bold projections that artificial intelligence will accelerate scientific discovery have raced ahead of evidence from working scientists, and the field

Context-Aware Deep Learning for Defect Classification in Atomic-Resolution STEM

AgentsDGX agent

arXiv:2606.09419v1 Announce Type: cross Abstract: Artificial intelligence is rapidly advancing materials characterization, yet most applications in electron microscopy rely solely on image contrast, o

Context-Fractured Decomposition Attacks on Tool-Using LLM Agents: Exploiting Artifact Provenance Gaps

SafetyDGX agent

arXiv:2606.09084v1 Announce Type: cross Abstract: Tool-using LLM agents interact with the world through actions that persist state in artifacts (e.g., workspace files or logs). Consequently, jailbreak

Context Over Compute Human-in-the-Loop Outperforms Iterative Chain-of-Thought Prompting in Interview Answer Quality

SafetyDGX agent

arXiv:2603.09995v2 Announce Type: replace-cross Abstract: Behavioral interview evaluation using large language models presents unique challenges that require structured assessment, realistic interview

Context Rot in AI-Assisted Software Development: Repurposing Documentation Consistency for AI Configuration Artifacts

Model ReleasesDGX agent

arXiv:2606.09090v1 Announce Type: cross Abstract: Developers increasingly provide AI coding assistants with persistent context through configuration files such as CLAUDE.md, AGENTS.md, and .cursorrule

Continual Quadruped Robots Coordination via Semantic Skill Discovery

AgentsDGX agent

arXiv:2606.08102v1 Announce Type: cross Abstract: Multi-quadruped coordination has attracted increasing attention due to its enhanced payload capacity, broader contact coverage, and improved adaptabil

Contract2Tool: Learning Preconditions and Effects for Reliable Tool-Augmented LLM Agents

AgentsDGX agent

arXiv:2606.07904v1 Announce Type: new Abstract: Tool-augmented large language model agents increasingly rely on external APIs, but standard tool schemas describe how to call a tool, not when the tool

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers

SafetyDGX agent

arXiv:2606.07604v1 Announce Type: cross Abstract: Analyzing attention weights has become a standard approach for interpreting the information flow of Large Language Models (LLMs). However, this approa

Correct Looks Better: Pairwise Comparisons Reveal Accuracy Rankings

SafetyDGX agent

arXiv:2606.09409v1 Announce Type: new Abstract: Pairwise comparisons combined with aggregation methods like Elo have become central to evaluating generative models, yet concerns remain that they rewar

Correcting Mean Bias in Text Embeddings: A Refined Renormalization with Training-Free Improvements on MMTEB

Model ReleasesDGX agent

arXiv:2511.11041v2 Announce Type: replace-cross Abstract: We find that current sentence-embedding models produce outputs with a consistent bias: every embedding e decomposes as ilde e + mu, where the

Correlation Is Not Enough: Embedding Human Metadata for Individual Causal Discovery

Model ReleasesDGX agent

arXiv:2606.09672v1 Announce Type: new Abstract: Ask a pretrained biomedical language model whether 'cortisol 28 ug/dL' and 'stock-market volatility' are related, and it returns a cosine similarity of

Cosmo3DFlow: Wavelet Flow Matching for Spatial-to-Spectral Compression in Reconstructing the Early Universe

ResearchDGX agent

arXiv:2602.10172v2 Announce Type: replace-cross Abstract: Reconstructing the early universe from the evolved present-day universe is a challenging and computationally demanding problem in modern astro

Cost-Aware Speculative Execution for LLM-Agent Workflows: An Integrated Five-Dimension Method

AgentsDGX agent

arXiv:2606.07846v1 Announce Type: cross Abstract: LLM-agent workflows chain model calls and tool invocations, and spend most of their wall-clock time waiting on upstream operations before downstream o

CoVEBench: Can Video Editing Models Handle Complex Instructions?

Model ReleasesDGX agent

arXiv:2606.08415v1 Announce Type: cross Abstract: While recent text-guided video editing models excel at elementary tasks (e.g., style transfer, object insertion), real-world user requests are highly

Crop Recommendation and Agricultural Query Answering System Using Spatio-Temporal Graph Neural Networks and Hybrid Retrieval Augmentation

ResearchDGX agent

arXiv:2606.09160v1 Announce Type: cross Abstract: This paper presents a unified system designed to support precision agriculture by integrating advanced weather prediction, crop recommendation, and a

Cross-LLM Consistency in Inference: Evidence from Shared Interactions

ResearchDGX agent

arXiv:2606.08129v1 Announce Type: new Abstract: Large language models (LLMs) differ in architecture, training data, and optimization procedures, yet they may still develop similar internal inference p

Cross-View Urban Traffic Dataset: Drone-Supervised Ground Truth for Monocular Bird's-Eye View Localization

Model ReleasesDGX agent

arXiv:2606.07708v1 Announce Type: cross Abstract: We introduce a dataset and benchmark for cross-view urban traffic perception built from synchronized ego-centric bicycle videos and aerial drone video

CT-VAM: A Cerebello-Thalamic-Inspired Vision-Action Model for Efficient Visuomotor Control

Local AiDGX agent

arXiv:2606.09572v1 Announce Type: cross Abstract: Vision-language-action models have shown strong promise for robot manipulation, yet raw language is primarily needed to specify task intent rather tha

Culturally-Adapted Red-Teaming Across East and Southeast Asian Contexts: A Methodological and Comparative Analysis

SafetyDGX agent

arXiv:2606.09178v1 Announce Type: cross Abstract: Multilingual safety evaluation of large language models (LLMs) has predominantly relied on direct translation (DT) of English benchmarks into target l

Curation of a Cardiology Interface Terminology for Highlighting Electronic Health Records using Machine Learning

ResearchDGX agent

arXiv:2606.08311v1 Announce Type: new Abstract: Electronic health record (EHR) notes are dense medical documents containing large amounts of information, often filled with complex medical jargon. High

CURE: Curriculum-guided Multi-task Training for Reliable Anatomy Grounded Report Generation

SafetyDGX agent

arXiv:2601.15408v2 Announce Type: replace-cross Abstract: Medical vision-language models can automate the generation of radiology reports but struggle with accurate visual grounding and factual consis

Customer Churn Prediction on Structured Data Using FT-Transformer and Stacking Ensembles

ResearchDGX agent

arXiv:2606.07582v1 Announce Type: cross Abstract: Customer churn prediction is essential across data-driven industries such as insurance, digital banking, eCommerce, and subscription platforms, where

Data Agents Under Attack: Vulnerabilities in LLM-Driven Analytical Systems

SafetyDGX agent

arXiv:2606.08661v1 Announce Type: cross Abstract: Data agents integrate LLM-driven reasoning with relational data access, executable analytical tools, and multi-step workflow orchestration, making the

Data Synthesis and Parameter-Efficient Fine-Tuning for Low-Resource NMT: A Case Study on Q'eqchi' Mayan

Model ReleasesDGX agent

arXiv:2606.09767v1 Announce Type: cross Abstract: Neural machine translation for digitally low-resource Indigenous languages is often hindered by extreme data scarcity, prompting reliance on extractiv

Dealing with Annotator Disagreement in Hate Speech Classification

ResearchDGX agent

arXiv:2502.08266v3 Announce Type: replace-cross Abstract: Hate speech detection is a crucial task, especially on social media where harmful content can spread quickly. Collecting social media content

Decision-Aware Memory Cards: Counterfactual-Inspired Context Selection and Compression for Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2606.08151v1 Announce Type: new Abstract: Tool-using LLM agents often fail not because relevant text is absent, but because decisive evidence is not selected, compressed, or surfaced at action t

Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models

Model ReleasesDGX agent

arXiv:2606.09142v1 Announce Type: cross Abstract: Egocentric vision offers a first-person view of human perception and decision making, yet its potential for traffic-safety prediction remains underexp

Decoupling Semantics and Logic: A Training-Free Coarse-to-Fine Pipeline for Video Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2606.07924v1 Announce Type: cross Abstract: This paper presents our system description for the 2nd Workshop on Multimodal Augmented Generation via MultimodAl Retrieval (MAGMaR). Addressing the c

Decoupling the 'What' and 'Where' With Polar Coordinate Positional Embeddings

ResearchDGX agent

arXiv:2509.10534v3 Announce Type: replace-cross Abstract: The attention mechanism in a Transformer architecture matches key to query based on both content -- the what -- and position in a sequence --

Deep Active Re-Labeling: Toward Noise-Resilient Annotation Efficiency

ResearchDGX agent

arXiv:2606.08718v1 Announce Type: cross Abstract: While Deep Active Learning (DAL) effectively reduces human annotation costs, its efficacy is constrained by human annotation errors. This is because t

Deep Tree Tensor Networks

Model ReleasesDGX agent

arXiv:2502.09928v2 Announce Type: replace-cross Abstract: Originating in quantum physics, tensor networks (TNs) have been widely adopted as exponential machines and parametric decomposers for recognit

Defending Against Malicious Finetuning by Scaling Train-time Adversarial Attacks

Model ReleasesDGX agent

arXiv:2606.07970v1 Announce Type: cross Abstract: Current open-weight large language models (LLMs) are prone to malicious finetuning attacks, which could compromise the safety alignment of LLMs with o

DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback

AgentsDGX agent

arXiv:2605.22781v2 Announce Type: replace-cross Abstract: LLM-powered AI agents require high-frequency state exploration (e.g., test-time tree search and reinforcement learning), relying on rapid chec

Deterministic Integrity Gates for LLM-Assisted Clinical Manuscript Preparation: An Auditable Biomedical Informatics Architecture

ResearchDGX agent

arXiv:2606.09500v1 Announce Type: new Abstract: Objective. Large language models (LLMs) increasingly draft clinical research manuscripts, but their fluency can hide fabricated citations, numbers that

Developing Distance-Aware Physics-Constrained Probabilistic Frameworks for Industrial Prognostics

Model ReleasesDGX agent

arXiv:2512.08499v3 Announce Type: replace-cross Abstract: Development of reliable and physically interpretable probabilistic frameworks for industrial prognostics remain nascent, and existing literatu

Difference-Aware Retrieval Policies for Imitation Learning

Local AiDGX agent

arXiv:2606.09758v1 Announce Type: cross Abstract: Parametric imitation learning via behavior cloning can suffer from poor generalization to out-of-distribution states due to compounding errors during

DiffoR: A Unified Continuous Generative Framework for Universal Ordinal Regression

ResearchDGX agent

arXiv:2606.07599v1 Announce Type: cross Abstract: Ordinal Regression (OR) aims to predict target values with inherent order, underpinning critical applications across diverse domains, from recommender

Discovering Data Structures: Nearest Neighbor Search and Beyond

Local AiDGX agent

arXiv:2411.03253v2 Announce Type: replace-cross Abstract: We propose a general framework for end-to-end learning of data structures. Our framework adapts to the underlying data distribution and provid

Discovering Expert-Level Nash Equilibrium Algorithms with Large Language Models

ResearchDGX agent

arXiv:2508.11874v2 Announce Type: replace-cross Abstract: Designing polynomial-time algorithms for approximate Nash equilibria (ANE) with provable worst-case guarantees is a fundamental open problem i

Discovering heuristics in a complex SAT solver with large language models

Model ReleasesDGX agent

arXiv:2507.22876v2 Announce Type: replace Abstract: The Satisfiability problem (SAT) is fundamental in computational complexity theory and has a wide range of industrial applications. Optimizing moder

Distilling LLM Reasoning into an Interpretable Policy Tree for Human-AI Collaboration

SafetyDGX agent

arXiv:2606.08596v1 Announce Type: new Abstract: Constructing efficient and reliable policies to assist humans is indispensable for human-AI collaboration. Existing methods mainly follow two lines of w

DIVERGE: Diversity-Enhanced RAG for Open-Ended Information Seeking

SafetyDGX agent

arXiv:2602.00238v2 Announce Type: replace-cross Abstract: Existing retrieval-augmented generation (RAG) systems often assume that each query has a single correct answer. This assumption overlooks open

Diverse Thinking Schemata Elicit Better Reasoning in Large Language Models

SafetyDGX agent

arXiv:2606.08974v1 Announce Type: new Abstract: Large reasoning models (LRMs) have attracted increasing attention for their ability to solve complex mathematical problems by generating extended reason

DIYHealth Suite: Dataset, Model, and Benchmark for Health Management at Home

Model ReleasesDGX agent

arXiv:2606.07542v1 Announce Type: cross Abstract: Generative AI is reshaping healthcare, yet most existing advances rely on hospital-grade devices, which limits their accessibility and potential for h

DN-Hypo-Pipeline: An AI-Driven Workflow for Hypothesis Generation via Large Language Models and Scientific Explanations

ResearchDGX agent

arXiv:2606.08532v1 Announce Type: new Abstract: A scientific hypothesis is the first step in research and undergoes experimental validation, yet it also reflects a deep understanding of and reasoning

Do Video Foundation Models Understand Intuitive Physics? A Layerwise Probing Analysis

ResearchDGX agent

arXiv:2606.09646v1 Announce Type: cross Abstract: We study whether pretrained video foundation models encode intuitive-physics information in their frozen representations, and how this information var

Does Persona Make LLMs K-pop Fans? A Pilot Study of LLM-Based Online Concert Audience Agents

SafetyDGX agent

arXiv:2606.07837v1 Announce Type: cross Abstract: A concert is a collective experience, but recorded performance videos are typically watched alone, stripping away the shared audience presence that ma

DOG-DPO:Dynamic Optimization in Geometry for Safety Alignment

SafetyDGX agent

arXiv:2606.07678v1 Announce Type: cross Abstract: Safety alignment for large language models relies on preference data, but current pipelines often train on large, redundant datasets. Existing data se

DOME: Learning Transferable Domain Variables from Sparse Supervision for Test-Time Adaptation

ApplicationsDGX agent

arXiv:2606.07646v1 Announce Type: cross Abstract: Test-time adaptation (TTA) aims to align a model to shifting test domains using only unlabeled streaming data. Most existing methods implicitly infer

DSFNet: Learning Dual-Domain Spectral Operators for Multi-Modality Spatio-Temporal Forecasting in Urban Transportation Systems

ApplicationsDGX agent

arXiv:2606.07695v1 Announce Type: cross Abstract: Multi-Modality Spatio-Temporal Forecasting (MoSTF) extends traditional spatio-temporal forecasting by incorporating diverse traffic modalities. Despit

Dynamic Distributed Constraint Optimization and Metareasoning for Continual, Large-Scale Satellite Operations

Local AiDGX agent

arXiv:2601.06188v3 Announce Type: replace Abstract: As Earth-observing satellite constellations grow in size and capability, distributed onboard control offers a pathway to novel responses and time-se

DynaOD: Dynamic Origin-Destination Flow Generation with Discrete-to-Continuous Temporal Semantic Modeling

ApplicationsDGX agent

arXiv:2606.09086v1 Announce Type: new Abstract: Dynamic origin-destination (OD) flow generation seeks to synthesize realistic mobility dynamics from temporal context alone, without relying on historic

EditSR: Enhancing Neural Symbolic Regression via Edit-based Rectification

TutorialsDGX agent

arXiv:2606.07915v1 Announce Type: new Abstract: Neural symbolic regression models improve inference efficiency by shifting structural search to pretraining, but their one-pass autoregressive decoding

Efficient Onboard Vision-Language Inference in UAV-Enabled Low-Altitude Economy Networks via LLM-Enhanced Optimization

ResearchDGX agent

arXiv:2510.10028v2 Announce Type: replace-cross Abstract: The rapid advancement of Low-Altitude Economy Networks (LAENets) has enabled a variety of applications, including aerial surveillance, environ

Efficient Skill Grounding via Code Refactoring with Small Language Models

Local AiDGX agent

arXiv:2606.07999v1 Announce Type: new Abstract: Effective skill grounding is essential for deploying reusable skills in embodied agents, as even minor embodiment or environmental differences can rende

Ego-Pi: VLA Fine-Tuning for Ego-Centric Human and Robot Data

TutorialsDGX agent

arXiv:2606.08107v1 Announce Type: cross Abstract: Robotics faces a fundamental challenge of data scarcity. Unlike language or vision research, there is no internet-scale dataset for robotic manipulati

← Previous
1…142143144145146…358
Next →