AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
30 Jun 2026

CMSL: Constructive Multi-Sequence Learning for Recommendation Systems

ResearchDGX agent

arXiv:2606.28533v1 Announce Type: cross Abstract: Sequence learning has emerged as the promising paradigm in recommendation systems, surpassing traditional Deep Learning Recommendation Models (DLRM) b

CMTFormer: Marrying Transformer with Hierarchical Information Interaction for RGB-Event Object Detection

SafetyDGX agent

arXiv:2606.29136v1 Announce Type: cross Abstract: Event cameras capture sparse brightness changes with high temporal resolution and high dynamic range, compensating for the deficiencies of the convent

Code Reasoning for Software Engineering Tasks: A Survey and A Call to Action

AgentsDGX agent

arXiv:2506.13932v3 Announce Type: replace-cross Abstract: The rise of large language models (LLMs) has led to dramatic improvements across a wide range of natural language tasks. Their performance on


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cognitive World Models for Process-Level Social Influence Evaluation

Model ReleasesDGX agent

arXiv:2606.29495v1 Announce Type: new Abstract: Social influence dialogue changes user behavior by altering internal cognitive states. The central evaluation question is whether the user's beliefs, de

COHORT: Collaborative Orchestration for Hardening via Offensive Replay on Emulated Topologies

Model ReleasesDGX agent

arXiv:2606.30479v1 Announce Type: cross Abstract: Mitigating an observed adversary in an enterprise network typically takes weeks of expert work: an analyst derives a mitigation tailored to that adver

Collective cooperation without individual fidelity in LLM agents

Model ReleasesDGX agent

arXiv:2606.30454v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as agents in simulations of social systems, yet it remains unclear when their behavior can be inter

ComMem: Complementary Memory Systems for Test-Time Adaptation of Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.28719v1 Announce Type: new Abstract: Test-time adaptation (TTA) of vision-language models (VLMs) is essential for their robust deployment in dynamic, real-world environments. However, exist

COMPASS: Grounding Composition-Intent Guidance in Unified Multimodal Models

ResearchDGX agent

arXiv:2606.28696v1 Announce Type: new Abstract: Composition is a high-level visual intent that governs where subjects are placed and how a scene is organized, yet current unified multimodal models rem

Compositional Dynamics in Learning and Mechanics

Model ReleasesDGX agent

arXiv:2606.28984v1 Announce Type: cross Abstract: We give a single compositional setting in which gradient-based learning and Hamiltonian-style mechanics appear as functorial semantics. The syntax is

ConCise: Training-Free Conclusion-Chain State Compression for Cost-Efficient Multi-Step RAG Services

HardwareDGX agent

arXiv:2606.28361v1 Announce Type: cross Abstract: Multi-step retrieval-augmented generation (RAG) has been widely deployed as LLM-powered web services for complex question answering, where iterative r

Confidence-feedback-weighted graph matching network: online-offline laser-induced damage site matching under complex interference

TutorialsDGX agent

arXiv:2606.29255v1 Announce Type: cross Abstract: Online inspection images of final optics in high-power laser facilities contain pseudo-damage sites that closely resemble true damage sites. Determini

Constrained Tabular Diffusion for Finance

ApplicationsDGX agent

arXiv:2606.28674v1 Announce Type: cross Abstract: Generative models in finance face the dual challenge of producing realistic data while satisfying strict regulatory and economic objectives, a require

Conversational Query Engine for Mixed-Modality Heterogeneous Enterprise Data Sources

Model ReleasesDGX agent

arXiv:2606.28370v1 Announce Type: cross Abstract: Enterprise business intelligence queries span structured warehouses and unstructured document repositories -- modalities with fundamentally different

Correct codes for the wrong reasons? validating LLMs as measurement instruments for theoretical constructs

ResearchDGX agent

arXiv:2606.28574v1 Announce Type: cross Abstract: When a large language model (LLM) codes a construct in text as a human annotator would, that agreement makes the LLM a reliable coder. Yet reliability

CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents

Model ReleasesDGX agent

arXiv:2511.02734v3 Announce Type: replace Abstract: Current evaluations of Large Language Model (LLM) agents primarily emphasize task completion, often overlooking resource efficiency and adaptability

Counterfactual Residual Data Augmentation for Regression

Model ReleasesDGX agent

arXiv:2606.28460v1 Announce Type: cross Abstract: Data-driven modeling in real-world regression tasks often suffers from limited training samples, high collection costs, and noisy observations. Inspir

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM

ResearchDGX agent

arXiv:2606.29563v1 Announce Type: cross Abstract: Large language models (LLMs) excel at complex tasks like question answering and summarization, thanks to their ability to handle long-context inputs.

Covering the Unseen: Information Demand Coverage Optimization for Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2606.29328v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) typically treats context selection as ranking chunks against a single query embedding. This assumption breaks dow

CRAFT: Counterfactual Credit Assignment from Free Sibling Rollouts for Self-Distilled Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.29476v1 Announce Type: cross Abstract: Self-distilled agentic reinforcement learning augments trajectory-level reward with a token-level distillation loss, using as its teacher the same pol

Critical Interval MSE: Toward Reliable Offline Validation for Robot Manipulation Policies

SafetyDGX agent

arXiv:2606.29898v1 Announce Type: cross Abstract: Real-world evaluation is the gold standard for robot policies because it tests them against the physical conditions and deployment challenges they are

Curvature-Guided Sheaf Diffusion for Unsupervised Community Detection on Heterophilic Graphs

ResearchDGX agent

arXiv:2606.30249v1 Announce Type: cross Abstract: Detecting communities in heterophilic graphs -- where connected nodes often belong to different classes -- is hard for unsupervised methods: classical

Customized Generative AI Agent for Transportation Engineering Practice: A Development and Continued Pre-training Guideline

Model ReleasesDGX agent

arXiv:2606.29014v1 Announce Type: new Abstract: Recent advancements in generative artificial intelligence (AI) and large language models (LLMs) have shown significant promise in automating complex rea

CW-B: Class Weighted Boosting Framework for Imbalance Resilient Multi Class Cardiac Phenotyping

ApplicationsDGX agent

arXiv:2606.29907v1 Announce Type: cross Abstract: Cardiac discharge phenotyping informs post-discharge treatment and follow-up, but real-world records are often incomplete and class-imbalanced, increa

CytoCLIP: Learning Cytoarchitectural Characteristics in Developing Human Brain Using Contrastive Language Image Pre-Training

TutorialsDGX agent

arXiv:2601.12282v2 Announce Type: replace-cross Abstract: The functions of different regions of the human brain are closely linked to their distinct cytoarchitecture, which is defined by the spatial a

DAPS++: Rethinking Diffusion Inverse Problems with Decoupled Posterior Annealing

TutorialsDGX agent

arXiv:2511.17038v4 Announce Type: replace Abstract: From a Bayesian perspective, score-based diffusion solves inverse problems through joint inference, embedding the likelihood with the prior to guide

Data and Evaluation Closed-Loop for Model Capability Enhancement

Model ReleasesDGX agent

arXiv:2606.28471v1 Announce Type: new Abstract: Model capability is the central variable in LLM pre-training, yet is never observed directly: data shapes it prospectively, while evaluation reveals it

Data-Efficient Multimodal Alignment for Histopathology-based Molecular Prediction

SafetyDGX agent

arXiv:2606.29949v1 Announce Type: cross Abstract: H&E-stained whole-slide images offer cohort-scale availability and rich spatial context but lack molecular specificity, whereas bulk RNA-seq provides

Data Provenance for Image Auto-Regressive Generation

ResearchDGX agent

arXiv:2606.28386v1 Announce Type: cross Abstract: Image autoregressive models (IARs) have recently demonstrated remarkable capabilities in visual content generation, achieving photorealistic quality a

Database Context Compression for Text-to-SQL on Real-World Large Databases

Model ReleasesDGX agent

arXiv:2606.28601v1 Announce Type: cross Abstract: Recent progress in Text-to-SQL has been driven by stronger language models and prompting strategies, yet performance on real enterprise benchmarks suc

Decomposing Memorization Reduction in Privacy-Preserving Fine-Tuning of SLMs for CSIRTs

ResearchDGX agent

arXiv:2606.28479v1 Announce Type: cross Abstract: CSIRTs increasingly fine tune language models on vulnerability scan records, but these records expose internal network topology and create privacy ris

DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective Verification

AgentsDGX agent

arXiv:2606.29746v1 Announce Type: new Abstract: Navigating the deluge of heterogeneous medical data, from academic literature (PubMed) to clinical guidelines (Web) and private knowledge bases, remains

DeepTrans Studio: Turning Expert Interventions into Shared Team Knowledge in Agentic Translation Workflows

AgentsDGX agent

arXiv:2606.29727v1 Announce Type: new Abstract: Professional translation is often a team-based process: translators, reviewers, and project managers must coordinate terminology, legal force, and accou

Defeat Devices in AI Systems

Model ReleasesDGX agent

arXiv:2606.28863v1 Announce Type: cross Abstract: AI systems increasingly exhibit behavior that differs systematically between evaluation and deployment contexts. Alignment faking, sandbagging, benchm

Defending Against Harmful Supervision Hidden in Benign Samples

ResearchDGX agent

arXiv:2606.30263v1 Announce Type: cross Abstract: Existing defenses are effective when harmful content is explicitly mixed into downstream fine-tuning data, but crafted samples can instead hide harmfu

Demonstration-Free Robotic Control via LLM Agents

Model ReleasesDGX agent

arXiv:2601.20334v2 Announce Type: replace-cross Abstract: Robotic manipulation has increasingly adopted vision-language-action (VLA) models, which achieve strong performance but typically require task

Deterministic Decisions for High-Stakes AI. A Zero-Egress Pipeline with the Deployability of RAG and the Accuracy of Machine Learning

SafetyDGX agent

arXiv:2606.29280v1 Announce Type: cross Abstract: We identify intervention bias as a previously unquantified failure mode of zero-shot large-language-model (LLM) educational advisory agents: without t

Diagnosing and Mitigating Context Rot in Long-horizon Search

ResearchDGX agent

arXiv:2606.29718v1 Announce Type: cross Abstract: Extensive context has become the norm as Large Language Models (LLMs) are increasingly deployed in long-horizon tasks. The concern that increasing con

Diagnosing and Repairing Factual Errors in RAG under Budget Constraints

Local AiDGX agent

arXiv:2606.29377v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) improves the factuality of large language models by grounding responses in external evidence, yet real-world deploy

Diff-Based Code Corruption using LLMs for Large-Scale Bugfix Benchmarking

Model ReleasesDGX agent

arXiv:2606.29088v1 Announce Type: cross Abstract: There are various benchmarks to evaluate bugfixing capabilities of Large Language Models. However, most widespread benchmarks do not fully reflect rea

Diff-MN: Diffusion Parameterized MoE-NCDE for Continuous Time Series Generation with Irregular Observations

TutorialsDGX agent

arXiv:2601.13534v3 Announce Type: replace-cross Abstract: Time series generation (TSG) is widely used across domains, yet most existing methods assume regular sampling and fixed output resolutions. Th

Digitizing Coaching Intelligence: An Agentic Framework for Holistic Athlete Profiling using VLM and RAG

Model ReleasesDGX agent

arXiv:2606.28570v1 Announce Type: cross Abstract: Athlete assessment is a critical process for tracking physical progress and identifying elite talent. However, during mass recruitment drives, traditi

Direct Causation in International Humanitarian Law and the Challenge of AI-Mediated Civilian Cyber Operations

AgentsDGX agent

arXiv:2606.29175v1 Announce Type: new Abstract: International humanitarian law protects civilians from direct attack unless and for such time as they take direct part in hostilities, with the ICRC's 2

DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning

Model ReleasesDGX agent

arXiv:2603.17863v2 Announce Type: replace-cross Abstract: Automating the development of machine learning algorithms has the potential to unlock new breakthroughs. However, our ability to improve and e

Distilling a Modular Reservoir Through a Genomic Bottleneck

TutorialsDGX agent

arXiv:2606.28380v1 Announce Type: cross Abstract: The intricate structures of biological neural networks largely emerge during development, guided by a comparatively compressed blueprint encoded in th

Distributionally Robust Reinforcement Learning with Human Feedback

SafetyDGX agent

arXiv:2503.00539v2 Announce Type: replace-cross Abstract: Reinforcement learning from human feedback (RLHF) has evolved to be one of the main methods for fine-tuning large language models (LLMs). Howe

Diversity is the Strength of the AI Crowd

Model ReleasesDGX agent

arXiv:2606.29661v1 Announce Type: new Abstract: Top AI forecasting systems are approaching superforecaster-level accuracy on future world events, but still rely primarily on off-the-shelf LLMs combine

DLR: Zero-Inference-Cost Latent Residuals for Low-Rank Pre-Training

Model ReleasesDGX agent

arXiv:2606.28932v1 Announce Type: cross Abstract: Large language models have driven recent progress in language and multimodal AI, yet pre-training them at scale is prohibitively expensive. Low-rank p

Do We Still Need Fine Tuning? Turkish Sentiment Analysis in the Era of Large Language Model

ResearchDGX agent

arXiv:2606.29614v1 Announce Type: cross Abstract: This study examines whether supervised fine-tuning remains necessary for Turkish sentiment analysis in the era of large language models. We compare cl

Dockerless: Environment-Free Program Verifier for Coding Agents

Model ReleasesDGX agent

arXiv:2606.28436v1 Announce Type: cross Abstract: Program verifiers play a central role in training coding agents, including selecting trajectories for supervised fine-tuning (SFT) and providing rewar

Does Role Specialization Matter for Explanation Faithfulness in Mixture-of-Experts?

ResearchDGX agent

arXiv:2606.29613v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) architectures have recently been extended with role-based mechanisms for interpretability. This is typically done by assignin

Does Verbose Chain-of-Thought Really Help? In-Distribution Evidence that Content, Not Length, Matters

Model ReleasesDGX agent

arXiv:2606.30128v1 Announce Type: new Abstract: Chain-of-thought (CoT) prompting improves LLM reasoning, but the source is contested: do the intermediate steps help because they carry useful semantic

Domain Adaptation with Adaptive Imagination for Visual Reinforcement Learning under Limited Target Data

ApplicationsDGX agent

arXiv:2606.30192v1 Announce Type: new Abstract: Sim-to-real transfer remains a major obstacle for reinforcement learning (RL), especially for vision-based control where image observations exacerbate t

Domain-Informed Multi-View Self-Distillation for Astronomical Light-Curve Representation Learning with JEPA

Model ReleasesDGX agent

arXiv:2606.28446v1 Announce Type: cross Abstract: Light curves describe temporal variations in the brightness of celestial objects. Learning robust representations of light curves is essential for lar

DOPD: Dual On-policy Distillation

SafetyDGX agent

arXiv:2606.30626v1 Announce Type: new Abstract: On-policy distillation (OPD) offers superior capacity transfer by supervising student-sampled trajectories with dense token-level signals. To furnish hi

DR-GS: Physically-Based Deformable and Relightable 2D Gaussians

Model ReleasesDGX agent

arXiv:2606.29379v1 Announce Type: cross Abstract: Gaussian splatting (GS) has garnered significant attention in VR/AR and digital content creation due to its explicit parameterization and efficient re

DRIFT: Difficulty Routing Self-DIstillation with Rhythm-Gated Exploration and Success BuFfer Training

SafetyDGX agent

arXiv:2606.30345v1 Announce Type: cross Abstract: Enabling large language models to achieve stable self-improvement without external expert supervision remains a central challenge in complex reasoning

Dual-Flow Reinforcement Learning with State-Aware Exploration

SafetyDGX agent

arXiv:2606.29820v1 Announce Type: cross Abstract: In complex continuous-control reinforcement learning tasks, multimodal optimal actions often coincide with uncertain, multimodal return distributions,

DuoMem: Towards Capable On-Device Memory Agents via Dual-Space Distillation

Model ReleasesDGX agent

arXiv:2606.29961v1 Announce Type: cross Abstract: Large Language Model (LLM)-based agents can solve complex procedural tasks by interacting with environments over multiple turns, but this ability typi

DyGnROLE: Asymmetric Pretraining for Edge Classification on Dynamic Graphs

SafetyDGX agent

arXiv:2602.23135v2 Announce Type: replace-cross Abstract: Edge classification on directed dynamic graphs requires modeling interactions between source and destination nodes exhibiting asymmetrical beh

Dynamic Parsing and Updating Natural Language Specification using VLMs for Robust Vision-Language Tracking

Model ReleasesDGX agent

arXiv:2606.29357v1 Announce Type: cross Abstract: Vision-language tracking guided by natural language specifications leverages high-level semantic cues of target objects to substantially boost trackin

← Previous
1…109110111112113…358
Next →