AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
8 Jun 2026

VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models

SafetyDGX agent

arXiv:2602.03160v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with the diverse spectrum of human values remains a central challenge: preference-based methods often fail to

Watch, Remember, Reason: Human-View Video Understanding with MLLMs

SafetyDGX agent

arXiv:2606.07433v1 Announce Type: cross Abstract: Video understanding is being rapidly transformed by multimodal large language models (MLLMs), as research moves from short clips to long, multimodal,

WAV: Multi-Resolution Block Residual Routing for Deep Decoder-Only Transformers

ResearchDGX agent

arXiv:2606.06564v1 Announce Type: cross Abstract: Residual connections are central to training deep Transformers, but standard PreNorm residual streams aggregate sublayer updates with fixed unit weigh


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

What Matters When Cotraining Robot Manipulation Policies on Everyday Human Videos?

SafetyDGX agent

arXiv:2606.06627v1 Announce Type: cross Abstract: Human video datasets used for cotraining robot manipulation policies largely consist of curated demonstrations where motions are orchestrated to resem

What Your Posts Reveal: A Benchmark and Agentic Framework for User-Level Privacy Leakage on Social Media

Model ReleasesDGX agent

arXiv:2606.06784v1 Announce Type: cross Abstract: Public social media posts can reveal private information through weak cues scattered across text, images, or metadata. Such leakage is often cumulativ

When Does Multi-Agent Collaboration Help? An Entropy Perspective

AgentsDGX agent

arXiv:2602.04234v6 Announce Type: cross Abstract: Multi-agent systems (MAS) have emerged as a prominent paradigm for leveraging large language models (LLMs) to tackle complex tasks. However, the mecha

When is 3D Worth It? A Resource-Performance Frontier for CNNs and Transformers in Lung CT

ResearchDGX agent

arXiv:2606.06950v1 Announce Type: cross Abstract: Three-dimensional models are widely assumed preferable for volumetric medical imaging, yet their practical value depends on whether performance gains

When Large Language Models Fail in Healthcare: Evaluating Sensitivity to Prompt Variations

Model ReleasesDGX agent

arXiv:2606.07237v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in healthcare for tasks such as clinical question answering, diagnosis support, and report summariz

Where Rectified Flows Leak: Characterising Membership Signals Along the Interpolation Path

ResearchDGX agent

arXiv:2606.07271v1 Announce Type: cross Abstract: Understanding what generative models retain from training data remains challenging, with implications for copyright and privacy. Beyond verbatim repro

Which Anatomy Matters Under Limited Labels? A Data-Efficient Anatomy-Aware Benchmark for Cardiac Pathology Prediction

Model ReleasesDGX agent

arXiv:2606.06509v1 Announce Type: cross Abstract: Numerous medical imaging problems must be solved under limited labels and constrained compute, yet it remains unclear whether performance gains are dr

Whisper Hallucination Detection and Mitigation via Hidden Representation Steering and Sparse AutoEncoders

ResearchDGX agent

arXiv:2606.07473v1 Announce Type: cross Abstract: Whisper, a widely adopted ASR model, is known to suffer from hallucinations - coherent transcriptions generated for non-speech audio entirely disconne

Workflow-to-Skill: Skill Creation via Routing-Workflow-Semantics-Attachments Decomposition

SafetyDGX agent

arXiv:2606.06893v1 Announce Type: new Abstract: Large language model agents increasingly rely on Skills to encode procedural knowledge, yet high-quality Skills remain costly to hand-write. This paper

Zero-Shot Embedding Drift Detection: A Lightweight Defense Against Prompt Injections in LLMs

Model ReleasesDGX agent

arXiv:2601.12359v1 Announce Type: cross Abstract: Prompt injection attacks have become an increasing vulnerability for LLM applications, where adversarial prompts exploit indirect input channels such

6 Jun 2026

2-Step Agent: A Framework for the Interaction of a Decision Maker with AI Decision Support

AgentsDGX agent

arXiv:2602.21889v2 Announce Type: replace Abstract: Predictions from ML models support human decision making in several fields, including high-stakes ones such as healthcare and the judiciary. Yet, we

A Cartography of Open Collaboration in Open Source AI: Mapping Practices, Motivations, and Governance in 14 Open Large Language Model Projects

ResearchDGX agent

arXiv:2509.25397v2 Announce Type: replace-cross Abstract: The proliferation of open large language models (LLMs) is fostering a vibrant ecosystem in artificial intelligence (AI). However, the methods

A Finite Certificate for the Positive n=9 Vasc Inequality

AgentsDGX agent

arXiv:2606.06136v1 Announce Type: cross Abstract: We prove the positive-real n=9 case of the Vasc cyclic inequality. The proof was obtained with human-guided assistance from the AI agent MechMath Agen

A Framework for Measuring Appropriate Reliance on Set-Valued AI Advice

ResearchDGX agent

arXiv:2606.06081v1 Announce Type: new Abstract: Appropriate reliance on AI advice has become a central research theme in human-AI collaboration. Existing frameworks have focused exclusively on point p

A Motivational Architecture for Conversational AGI

AgentsDGX agent

arXiv:2606.05411v1 Announce Type: new Abstract: Motivational architectures in cognitive AI have largely been designed for physical agents regulating bodily needs. Conversational agents operate in a di

A Pre-Registered Causal Partition of Self-Consistency Elicitation and Reward Design in RLVR

SafetyDGX agent

arXiv:2606.05932v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) improves reasoning even when the reward signal is spurious -- assigning credit to the group-plural

A Taxonomy of Runtime Faults in Model Context Protocol Servers

AgentsDGX agent

arXiv:2606.05339v1 Announce Type: cross Abstract: MCP (Model Context Protocol) enables LLMs (Large Language Models) to interact with external tools and data sources via a standardized protocol. Its ra

A2RAG: Adaptive Agentic Graph Retrieval for Cost-Aware and Reliable Reasoning

AgentsDGX agent

arXiv:2601.21162v2 Announce Type: replace-cross Abstract: Graph Retrieval-Augmented Generation (Graph-RAG) enhances multihop question answering by organizing corpora into knowledge graphs and routing

AdaMEM: Test-Time Adaptive Memory for Language Agents

SafetyDGX agent

arXiv:2606.05684v1 Announce Type: new Abstract: A central challenge for language agents is utilizing past experience to adapt to dynamic test-time conditions. While recent work demonstrates the promis

Adapting Diffusion Language Models for Lossless Pixel-Level Image Transmission

ResearchDGX agent

arXiv:2606.06273v1 Announce Type: cross Abstract: Lossless pixel-level image transmission is a fundamental regime beyond semantic communications, because exact recovery requires both accurate symbol p

ADK Arena: Evaluating Agent Development Kits via LLM-as-a-Developer

Model ReleasesDGX agent

arXiv:2606.05548v1 Announce Type: cross Abstract: The rapid proliferation of Agent Development Kits (ADKs), SDK-level frameworks for building LLM-powered autonomous agents, has outpaced any empirical

Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning

AgentsDGX agent

arXiv:2503.01734v3 Announce Type: replace-cross Abstract: Attacks on machine learning models have been extensively studied through stateless optimization. In this paper, we demonstrate how a reinforce

Agent Memory: Characterization and System Implications of Stateful Long-Horizon Workloads

Model ReleasesDGX agent

arXiv:2606.06448v1 Announce Type: new Abstract: LLM agents are increasingly deployed on long-horizon tasks requiring sustained reasoning over extended interaction histories. Realizing this at scale re

Agent-Orchestrated Adaptive RAG: A Comparative Study on Structured and Multi-Hop Retrieval

Model ReleasesDGX agent

arXiv:2606.05658v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by grounding their responses in external knowledge, but conventional pipeli

Agentic Molecular Recovery via Molecule-Aware Exploration

AgentsDGX agent

arXiv:2606.05847v1 Announce Type: new Abstract: Text-guided molecular generation with LLMs often yields invalid SMILES. We argue that invalid drafts should be addressed through a shift from validity-o

Agentic Monte Carlo: Simulating Reinforcement Learning for Black-Box Agents

Model ReleasesDGX agent

arXiv:2606.05296v1 Announce Type: cross Abstract: LLM agents operate in two distinct regimes: open-weight agents amenable to reinforcement learning (RL) and black-box agents whose behaviour must be co

AIS-Based Vessel Trajectory Prediction Using Memory-Augmented Neural Networks

ResearchDGX agent

arXiv:2606.06311v1 Announce Type: new Abstract: Accurate vessel trajectory prediction is essential for safe and efficient maritime operations, enabling collision avoidance and supporting route optimiz

Amortizing Federated Adaptation: Hypernetwork Driven LoRA for Personalized Foundation Models

SafetyDGX agent

arXiv:2606.06154v1 Announce Type: new Abstract: Federated fine-tuning of foundation models using Low-Rank Adaptation (LoRA) offers a communication efficient solution for distributed learning. However,

An Improved CNN-LSTM Based Intrusion Detection System for IoT Networks

ResearchDGX agent

arXiv:2606.05776v1 Announce Type: cross Abstract: With the rapid proliferation of IoT devices, security concerns have dramatically escalated and intrusion detection systems have become critical for pr

An Infectious Disease Spread Simulation Based on Large Language Model Decision Making

SafetyDGX agent

arXiv:2606.06360v1 Announce Type: new Abstract: Modelling individual decision-making during infectious disease outbreaks is crucial for understanding behavioural dynamics and informing effective publi

An interpretable and trustworthy AI framework for large-scale longitudinal structure-pain association studies using data from the Osteoarthritis Initiative (OAI)

ResearchDGX agent

arXiv:2606.05357v1 Announce Type: new Abstract: Purpose: To develop an interpretable and trustworthy AI framework that combines deep learning based MRI Osteoarthritis Knee Score (MOAKS) prediction wit

Answer Presence Drives RAG Rewriting Gains

ResearchDGX agent

arXiv:2606.05633v1 Announce Type: new Abstract: Retrieval-augmented QA pipelines often route retrieved passages through an LLM rewriter before a smaller reader, lifting F1 by tens of points on multi-h

Assessing the Carbon Emissions and Energy Consumption of U.S. Hyperscale Data Centers

ResearchDGX agent

arXiv:2606.05420v1 Announce Type: new Abstract: The rapid proliferation of hyperscale data centers (HDCs) in the US, mainly driven by the adoption of artificial intelligence, has raised concerns about

Assessing the Geographic Diversity of AI's Platial Representations in Image Generation

SafetyDGX agent

arXiv:2606.05188v1 Announce Type: cross Abstract: (Gen)AI diversity is not merely an ethical issue. From the perspective of geographic information science (GIScience), it could be interpreted as a fun

AttackPathGNN: Cross-function vulnerability detection in smart contracts using state interference graphs and conjunction pooling

Model ReleasesDGX agent

arXiv:2606.05986v1 Announce Type: cross Abstract: Existing learning-based detectors for Solidity smart-contracts reduce vulnerability detection to syntactic pattern matching within single functions, y

Balancing Image Compression and Generation with Bootstrapped Tokenization

Local AiDGX agent

arXiv:2606.05552v1 Announce Type: cross Abstract: Despite progress in image tokenization, standard methods encode redundant information by mixing all granularities within each token, thus redundancy p

Benchmark Everything Everywhere All at Once

Model ReleasesDGX agent

arXiv:2606.06462v1 Announce Type: new Abstract: Benchmarks are fundamental for evaluating and advancing LLMs and MLLMs by providing standardized and explicit measures of performance. However, their co

Benchmarking Counterfactual Prediction in Epidemic Time Series with Time-Varying Interventions

Model ReleasesDGX agent

arXiv:2606.05692v1 Announce Type: cross Abstract: Deep learning has enabled significant advances in time-series causal inference, yet progress remains constrained by the lack of realistic benchmarks w

Benchmarks in Leipzig

ResearchDGX agent

arXiv:2606.05818v1 Announce Type: cross Abstract: Between April 1 and May 15, 2026, a group of 49 mathematicians compiled a dataset of research-level mathematics questions with known answers. Most of

Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation

HardwareDGX agent

arXiv:2512.03086v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown remarkable capabilities in code translation, yet their performance deteriorates in low-resource progra

Beyond Means: Topological Causal Effects under Persistent-Homology Ignorability

ResearchDGX agent

arXiv:2603.14169v2 Announce Type: replace-cross Abstract: Average treatment effects (ATE) and conditional average treatment effects (CATE) are foundational causal estimands, but they target changes in

Beyond Output Matching: Preserving Internal Geometry in NVFP4 LLM Distillatio

Model ReleasesDGX agent

arXiv:2606.05682v1 Announce Type: new Abstract: Demand for low-precision inference, including NVFP4-based approaches, has grown as large language models are increasingly deployed in latency and cost c

Beyond Rewards in Reinforcement Learning for Cyber Defence

SafetyDGX agent

arXiv:2602.04809v3 Announce Type: replace-cross Abstract: Recent years have seen an explosion of interest in autonomous cyber defence agents trained to defend computer networks using deep reinforcemen

Beyond Semantic Organization: Memory as Execution State Management for Long-Horizon Agents

AgentsDGX agent

arXiv:2606.06090v1 Announce Type: new Abstract: LLM-based agents increasingly tackle long-horizon tasks with interdependent decisions, where each action reshapes future constraints and intermediate er

Beyond Similarity: Trustworthy Memory Search for Personal AI Agents

AgentsDGX agent

arXiv:2606.06054v1 Announce Type: new Abstract: Personal AI agents increasingly rely on long-term memory to provide persistent personalization across sessions. However, existing memory pipelines are l

Beyond Soft Masks: Hard-Perturbation Mixup Explainer for Robust GNN Explainability

ApplicationsDGX agent

arXiv:2606.05756v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) have demonstrated remarkable performance across a range of applications involving graph-structured data, particularly in

Beyond Vector Similarity: A Structural Analysis of Graph-Augmented Retrieval for Industrial Knowledge Graphs

ResearchDGX agent

arXiv:2606.06003v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) fails systematically on queries requiring structural reasoning over interconnected entities. We compare eight retri

Beyond Waveform Robustness: Robust Feature-Vocoder Adversarial Attacks on Automatic Speech Recognition

ResearchDGX agent

arXiv:2606.05678v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) systems have become widely used for multilingual speech-to-text transcription. Their robustness to adversarial atta

Bidirectional Search for Longest Paths: Case for Front-to-Front Heuristics

ResearchDGX agent

arXiv:2606.05956v1 Announce Type: new Abstract: Bidirectional heuristic search can potentially reduce search effort for problems amenable to backward search. Therein, it is well-known that front-to-fr

Binary Gaussian Copula Synthesis: an LLM-powered data augmentation framework for early dialysis prediction in chronic kidney disease

ApplicationsDGX agent

arXiv:2403.00965v2 Announce Type: replace-cross Abstract: Only a small fraction of patients with chronic kidney disease (CKD) progress to dialysis, creating severe class imbalance that limits the perf

Boosting Brain-to-Image Decoding with TRIBE v2 Data Augmentation

ResearchDGX agent

arXiv:2606.06345v1 Announce Type: new Abstract: Brain decoding is limited by the availability of labeled neural data, and remains challenging in low-data regimes. To address this issue, we investigate

Breaking the Chain: A Causal Analysis of LLM Faithfulness to Intermediate Structures

ResearchDGX agent

arXiv:2603.16475v2 Announce Type: replace Abstract: In schema-guided reasoning (SGR) pipelines, LLMs produce explicit intermediate structures -- rubrics, checklists, or verification queries -- before

Brick-Composer: Using MLLMs for Assembly with Diverse Bricks

Model ReleasesDGX agent

arXiv:2606.05445v1 Announce Type: new Abstract: We dream of AI agents that can read arbitrary designs and construct real-world objects from reusable building blocks. As a first step toward this vision

Bridging Domain Expertise and Generalization for Performance Estimation

SafetyDGX agent

arXiv:2606.06335v1 Announce Type: cross Abstract: Performance estimation under distribution shift aims to predict how a model behaves on an unlabeled test set whose distribution differs from the train

Bridging the Semantic-Collaborative Gap: An Asymmetric Graph Architecture for Cold-Start Item Recommendation

ApplicationsDGX agent

arXiv:2606.06225v1 Announce Type: cross Abstract: Collaborative filtering and graph-based recommendation models are highly effective because they leverage observed user interactions, but this dependen

CaMeLs Can Use Computers Too: System-level Security for Computer Use Agents

AgentsDGX agent

arXiv:2601.09923v3 Announce Type: replace Abstract: AI agents are vulnerable to prompt injection attacks, where malicious content hijacks agent behavior. Among proposed defenses, architectural isolati

Can AI Refute Economic Theory? Evidence from Beyond the Knowledge Cutoff

Model ReleasesDGX agent

arXiv:2606.05383v1 Announce Type: cross Abstract: Can artificial intelligence (AI) refute economic theory? I document experiments in which I asked several AI models (Gemini, Refine, Claude, and ChatGP

← Previous
1…154155156157158…358
Next →