AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
2 Jun 2026

Authenticity Debt and the Synthetic Content Threat Landscape: A Layered Framework for Trust, Provenance, and IP Governance in the Generative AI Era

SafetyDGX agent

arXiv:2606.00621v1 Announce Type: cross Abstract: Generative artificial intelligence has fundamentally changed how content is now produced. It has enabled how high-fidelity text, images, audio, and vi

AutoEval Done Right: Using Synthetic Data for Model Evaluation

Model ReleasesDGX agent

arXiv:2403.07008v3 Announce Type: replace-cross Abstract: The evaluation of machine learning models using human-labeled validation data can be expensive and time-consuming. AI-labeled synthetic data c

AutoForest: Automatically Generating Forest Plots from Biomedical Studies with End-to-End Evidence Extraction and Synthesis

Applications

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.02403v1 Announce Type: cross Abstract: Systematic reviews rely on forest plots to synthesise quantitative evidence across biomedical studies, but generating them remains a fragmented and la

Automated Conjecture Resolution with Formal Verification

AgentsDGX agent

arXiv:2604.03789v2 Announce Type: replace-cross Abstract: Recent advances in large language models have significantly improved their ability to perform mathematical reasoning, extending from elementar

Automatically Differentiable Nonlinear Tensor Networks (ADNTNs) for Exponential Compression of Deep Neural Networks

ResearchDGX agent

arXiv:2606.00130v1 Announce Type: cross Abstract: We study Automatically Differentiable Nonlinear Tensor Networks (ADNTNs), a family of structured weight generators whose compact core tensors are trai

AutoMedBench: Towards Medical AutoResearch with Agentic AI Models

Model ReleasesDGX agent

arXiv:2606.01961v1 Announce Type: new Abstract: Autonomous agents are increasingly expected to support end-to-end medical-AI research workflows, moving beyond isolated prediction tasks or short-form c

Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation

ResearchDGX agent

arXiv:2601.00664v2 Announce Type: replace-cross Abstract: Talking head generation creates lifelike avatars from static portraits for virtual communication and content creation. However, current models

AXIOM: A Trust-First Neuro-Symbolic Execution Architecture for Verifiable Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2606.00671v1 Announce Type: new Abstract: We present AXIOM, a trust-first neuro-symbolic execution architecture for natural-language mathematical reasoning. In AXIOM, the language model function

BADGER: Bridging Agentic and Deterministic Evaluation for Generative Enterprise Reasoning

Model ReleasesDGX agent

arXiv:2606.02109v1 Announce Type: new Abstract: Enterprise AI systems that translate natural language into SQL queries and orchestrate multi-step agentic reasoning pipelines require evaluation approac

BAGEN: Are LLM Agents Budget-Aware?

AgentsDGX agent

arXiv:2606.00198v1 Announce Type: cross Abstract: While agents are increasingly spending more resources, today agent cost is mostly measured only after execution. A Budget-Aware Agent (BAGEN) should t

Bayesian Inference of Nonlinear Malaria Dynamics in Ghana via an Ensemble Markov Chain Monte Carlo Sampler

Model ReleasesDGX agent

arXiv:2606.00783v1 Announce Type: cross Abstract: Reliable quantification of malaria dynamics in sub-Saharan Africa is hindered by short, noisy, and spatially heterogeneous surveillance records. In Gh

Bayesian Spectral Emotion Transition Discovery from Multi-Annotator Disagreement

ResearchDGX agent

arXiv:2606.01906v1 Announce Type: new Abstract: Emotions evolve through the dynamics of conversation, and understanding their transition structure is foundational to applications ranging from mental-h

Before the Model Learns the Bug:Fuzzing RLVR Verifiers

TutorialsDGX agent

arXiv:2606.01066v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) replaces human preference labels with executable reward functions such as math answer checkers, JS

Behavior-Invariant Task Representation Learning with Transformer-based World Models for Offline Meta-Reinforcement Learning

SafetyDGX agent

arXiv:2606.00780v1 Announce Type: cross Abstract: Offline meta-reinforcement learning leverages static datasets to enable agents to generalize to unseen environments by combining offline efficiency wi

BenchEvolver: Frontier Task Synthesis via Solution-Centric Evolution

Model ReleasesDGX agent

arXiv:2606.01286v1 Announce Type: cross Abstract: The rapid progress of frontier large language models has led to widespread benchmark saturation, limiting the ability of existing datasets to differen

Benchmarking Multimodal LLMs on Code Generation for Complex Interactive Webpages

Model ReleasesDGX agent

arXiv:2606.00154v1 Announce Type: cross Abstract: Recent advancements in multimodal large language models (MLLMs) have achieved remarkable progress in multimodal reasoning and code generation, catalyz

Benchmarking Security Risk Detection and Verification in Open Agentic Skill Ecosystems

Model ReleasesDGX agent

arXiv:2606.00925v1 Announce Type: cross Abstract: Open agent platforms allow community contributors to publish reusable skills that agents can invoke at runtime. This extensibility also creates a supp

Benchmarks for Vision-Language Models in Urban Perception Should Be Reliability-Aware and Negotiated

Model ReleasesDGX agent

arXiv:2606.00871v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used to generate structured descriptions of street-level imagery for tasks such as streetscape auditing

Better Source, Better Flow: Learning Condition-Dependent Source Distribution for Flow Matching

SafetyDGX agent

arXiv:2602.05951v2 Announce Type: replace-cross Abstract: Flow matching has recently emerged as a promising alternative to diffusion-based generative models, particularly for text-to-image generation.

Beware of the Batch Size: Hyperparameter Bias in Evaluating LoRA

Model ReleasesDGX agent

arXiv:2602.09492v2 Announce Type: replace-cross Abstract: Low-rank adaptation (LoRA) is a standard approach for fine-tuning large language models, yet its many variants report conflicting empirical ga

Beyond Access: Guided LLM Scaffolding for Independent Learning in Undergraduate Statistics

SafetyDGX agent

arXiv:2606.01375v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly entering students' learning practices, but their educational value depends on whether they support reaso

Beyond Augmentation: Score-Guided Pathological Prior for EEG-based Depression Detection

ResearchDGX agent

arXiv:2606.00180v1 Announce Type: cross Abstract: Deep learning-based Major Depressive Disorder (MDD) detection using Electroencephalography (EEG) is fundamentally constrained by the 'small-sample dil

Beyond Categories of Caste: Examining Caste Bias and Morality in Text-to-Image AI Models

SafetyDGX agent

arXiv:2606.00039v1 Announce Type: cross Abstract: Text-to-Image (T2I) models have shown promising utility across various domains. However, such models are also amplifying harmful societal biases in th

Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation

SafetyDGX agent

arXiv:2602.11790v2 Announce Type: replace Abstract: Although recent end-to-end video generation models demonstrate impressive performance in visually oriented content creation, they remain limited in

Beyond Independent Manipulation: Individual Fairness-aware Strategic Classification with Peer Imitation

SafetyDGX agent

arXiv:2606.00827v1 Announce Type: cross Abstract: Strategic classification (SC) investigates scenarios where agents manipulate their features to obtain favorable decisions from predictive models. Exis

Beyond Model Base Retrieval: Weaving Knowledge to Master Fine-grained Neural Network Design

ResearchDGX agent

arXiv:2507.15336v3 Announce Type: replace-cross Abstract: Designing high-performance neural networks for new tasks requires balancing optimization quality with search efficiency. Current methods fail

Beyond One-shot: AI Agents for Learning in Field Experiments

AgentsDGX agent

arXiv:2606.02458v1 Announce Type: new Abstract: Organizations routinely run experiments for A/B testing, yet the data generated from one experiment is underutilized to inform subsequent intervention d

Beyond String Matching: Semantic Evaluation of PDF Table Extraction

ResearchDGX agent

arXiv:2603.18652v2 Announce Type: replace-cross Abstract: Reliably extracting tables from PDFs is essential for large-scale scientific data mining and knowledge base construction, yet existing evaluat

Beyond Task-Agnostic: Task-Aware Grouping for Communication-Efficient Multi-Task MoE Inference

Local AiDGX agent

arXiv:2606.01007v1 Announce Type: cross Abstract: Sparsely activated Mixture-of-Experts (MoE) models scale capacity via conditional computation, but distributed inference suffers from cross-GPU expert

Beyond Task Success: Behavioral and Representational Diagnostics for WAM and VLA

ResearchDGX agent

arXiv:2606.01095v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies and World-Action Models (WAM) represent two increasingly important paradigms for robotic manipulation. However,

Beyond Text and Tables: Vision-Language Model Integration in ComProScanner for Extracting Materials Data from Scientific Figures with High Accuracy

Model ReleasesDGX agent

arXiv:2606.00065v1 Announce Type: cross Abstract: Automated extraction of materials composition-property data from scientific literature has advanced considerably with the development of large languag

Beyond the Mouth: Upper-Face Affective Cues in Audiovisual Sentence Recognition under Acoustic Uncertainty

ResearchDGX agent

arXiv:2606.00670v1 Announce Type: cross Abstract: Face-to-face speech comprehension is inherently multimodal, integrating acoustic signals with visible articulation, facial expression, head motion, an

Beyond Tool Adoption: A Practical Five-Stage Developmental Continuum for AI Literacy in Higher Education

TutorialsDGX agent

arXiv:2606.00038v1 Announce Type: cross Abstract: Artificial intelligence (AI) literacy is increasingly recognized as a foundational competency for all university graduates. Yet students' engagement w

Beyond Visual Memory: Mechanistic Diagnostics of Latent Visual Reasoning

ResearchDGX agent

arXiv:2606.01287v1 Announce Type: cross Abstract: Recent latent visual reasoning methods achieve substantial gains by inserting continuous latent tokens into multimodal language models. These gains ar

BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization

ResearchDGX agent

arXiv:2606.00079v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) large language models reduce per-token computation through sparse expert activation, but their deployment remains memory-inte

Boosting Multimodal Federated Learning via Chained Modality Optimization

ResearchDGX agent

arXiv:2606.01856v1 Announce Type: cross Abstract: Multimodal Federated Learning (MMFL) enables privacy-preserving collaborative learning across decentralized clients with heterogeneous data and modali

Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention

Model ReleasesDGX agent

arXiv:2512.10414v2 Announce Type: replace Abstract: Recently, reinforcement learning (RL) has become a common choice in enhancing the reasoning capabilities of vision-language models (VLMs). Consideri

Brain-Atlas-Guided Generative Counterfactual Attention for Explainable Cognitive Decline Diagnosis Using Multimodal Connectomes

ResearchDGX agent

arXiv:2606.01237v1 Announce Type: new Abstract: Mild cognitive impairment (MCI) and subjective cognitive decline (SCD) are closely associated with the early Alzheimer's disease continuum, where accura

Breaking the Information Silo: Semantic Personas for Cross-Domain Recommendation

SafetyDGX agent

arXiv:2606.01783v1 Announce Type: cross Abstract: Digital platforms increasingly operate as isolated information silos, limiting their ability to construct comprehensive user representations across do

Breaking the Reversal Curse in Autoregressive Language Models via Identity Bridge

SafetyDGX agent

arXiv:2602.02470v2 Announce Type: replace Abstract: Autoregressive large language models (LLMs) have achieved remarkable success in many complex tasks, yet they can still fail in very simple logical r

Bridging Reasoning Trajectories in On-Policy Distillation via Near-Future Guidance

SafetyDGX agent

arXiv:2606.00305v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) improves large language model reasoning by training a student model on trajectories sampled from its own policy under tea

Bridging Requirements and Architecture: Multi-Agent Orchestration with External Knowledge and Hierarchical Memory

Model ReleasesDGX agent

arXiv:2606.01385v1 Announce Type: cross Abstract: Software architecture design is a critical yet inherently complex and knowledge-intensive phase that requires balancing competing quality attributes a

Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation

ResearchDGX agent

arXiv:2606.00095v1 Announce Type: cross Abstract: Vision-Language Navigation (VLN) enables embodied agents to reach target locations in unseen environments by following language instructions. Despite

Bridging the Last Mile of Time Series Forecasting with LLM Agents

SafetyDGX agent

arXiv:2606.02497v1 Announce Type: new Abstract: Time series forecasting has advanced rapidly, especially with the emergence of foundation models that show strong zero-shot performance on numerical ext

Bridging the Sim-to-Real Gap in Semiconductor Visual Program Synthesis via Input Binarization

Model ReleasesDGX agent

arXiv:2606.02434v1 Announce Type: new Abstract: Precise parametric control over circuit geometry is essential for semiconductor inspection, yet obtaining sufficient real training data remains costly.

BRo-JEPA: Learning Modular Arithmetic in Latent Space

TutorialsDGX agent

arXiv:2606.01372v1 Announce Type: cross Abstract: Can neural networks learn abstract algebraic rules, or do they merely memorize training patterns? We investigate this using MNIST digits as states and

BudgetDraft: Acceptance-Aware Multi-View Training for Sparse-KV Speculative Decoding

HardwareDGX agent

arXiv:2606.00144v1 Announce Type: cross Abstract: Speculative decoding speeds up autoregressive decoding by using a drafter to propose multiple tokens that a verifier validates in parallel. In resourc

Business Utility of Large Language Models as Exploratory Data Analysis Agents

Model ReleasesDGX agent

arXiv:2606.00051v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in analytical workflows, but their suitability as exploratory data analysis (EDA) agents in busines

c-TPE: Tree-structured Parzen Estimator with Inequality Constraints for Expensive Hyperparameter Optimization

ApplicationsDGX agent

arXiv:2211.14411v5 Announce Type: replace-cross Abstract: Hyperparameter optimization (HPO) is crucial for strong performance of deep learning algorithms and real-world applications often impose some

CA-BED: Conversation-Aware Bayesian Experimental Design

ResearchDGX agent

arXiv:2606.01182v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at static reasoning tasks, yet their performance often degrades in interactive scenarios where information must be

CAFOSat: A Strongly Annotated Dataset for Infrastructure-Aware CAFO Mapping Using High-Resolution Imagery

Model ReleasesDGX agent

arXiv:2606.00548v1 Announce Type: cross Abstract: Concentrated Animal Feeding Operations (CAFOs) play an important role in agricultural production but are also associated with environmental, public he

Calibrating Uncertainty for Zero-Shot Adversarial CLIP

SafetyDGX agent

arXiv:2512.12997v2 Announce Type: replace-cross Abstract: CLIP delivers strong zero-shot classification but remains highly vulnerable to adversarial attacks. Prior adversarial fine-tuning work primari

Can AI Review Improve Paper Drafting? An Empirical Study on 20 Computer Architecture Submissions

SafetyDGX agent

arXiv:2606.01013v1 Announce Type: new Abstract: Research is advancing faster than ever with artificial intelligence (AI); and so are the corresponding research papers. The exploding volume of AI-gener

Can LLM Agents Sustain Long-Horizon Organizational Dynamics?

AgentsDGX agent

arXiv:2606.01199v1 Announce Type: new Abstract: Large language agents are increasingly used for social simulation, yet it remains unclear whether they can sustain coherent behavior in structured organ

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures

Model ReleasesDGX agent

arXiv:2505.24069v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are deployed on increasingly complex tasks that require multi-step decision-making. Understanding their algorithm

Can Predicted Dynamics Exist in the Physical World?

ResearchDGX agent

arXiv:2606.00089v1 Announce Type: cross Abstract: Predictive Physical AI systems output state rollouts, action chunks, and latent plans, yet a low root-mean-square error (RMSE) does not imply that a p

Capability Self-Assessment: Teaching LLMs to Know Their Limits

Local AiDGX agent

arXiv:2606.00251v1 Announce Type: new Abstract: The ability to recognize one's own limitations and decide whether to solve a problem or delegate is fundamental for reliable intelligent systems. Yet we

CAPF: Guiding Search-Agent Rollouts with Credit-Attenuated Privileged Feedback

SafetyDGX agent

arXiv:2606.01830v1 Announce Type: new Abstract: Recent LLM search agents use reinforcement learning with verifiable rewards (RLVR) to learn search-augmented reasoning from outcome rewards. On hard pro

CardioLens: Revealing the Clinical Reality Gap of MLLMs via Multi-Sequence Cardiac MRI Evaluations

ApplicationsDGX agent

arXiv:2606.00123v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown strong performance on public medical benchmarks, yet existing evaluations often remain weak proxie

CARE-RL: Capability-Aware Reinforcement Learning for Mitigating Cross-Domain Conflicts

ResearchDGX agent

arXiv:2606.00609v1 Announce Type: cross Abstract: Reinforcement learning (RL) with verifiable rewards has achieved strong progress in reasoning-oriented LLMs, but extending it to multi-domain RL remai

← Previous
1…170171172173174…358
Next →