AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
11 Aug 2026

Dual-Adversarial Safety Alignment: Cultivating Intrinsic Threat Comprehension in LRMs

SafetyDGX agent

arXiv:2608.09542v1 Announce Type: cross Abstract: Large reasoning models (LRMs) achieve remarkable success on complex tasks but remain vulnerable to harmful prompts that induce unsafe outputs. Recent

DualCert: A Solver for the Traveling Salesman Problem with Constraint-Coupled Learning

Model ReleasesDGX agent

arXiv:2608.09042v1 Announce Type: new Abstract: Large traveling salesman problem (TSP) instances require a solver to allocate limited computation while preserving the validity of its outputs. Existing

DUET: A Diversity-Quality Duet of Distillation Experts for Two-Step Video Generation

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.09637v1 Announce Type: cross Abstract: Diffusion models have enabled high-quality video generation in recent years, but the high cost of iterative sampling hinders their practical deploymen

Dynamic Coalition Formation and Communication Pricing in Skill-Based Agentic AI Systems

AgentsDGX agent

arXiv:2608.07532v1 Announce Type: new Abstract: Modern agentic AI systems combine multiple large language model agents with heterogeneous skills, yet most architectures either fix communication in adv

Dynamic gain neuromodulation attenuates the stability gap under joint training

ResearchDGX agent

arXiv:2507.14056v3 Announce Type: replace-cross Abstract: Recent work in continual learning has highlighted the stability gap -- a temporary performance drop on previously learned tasks when new ones

EasyBalance: Cross-Layer Load Balancing in Distributed MoE Inference

HardwareDGX agent

arXiv:2608.07964v1 Announce Type: cross Abstract: Load Balancing has emerged as a critical problem in expert-parallel distributed inference of Mixture-of-Experts (MoE) models. As routing distributions

Eco-SoC: A Sustainable VLSI Architecture for Energy-Proportional Artificial Intelligence

HardwareDGX agent

arXiv:2608.08761v1 Announce Type: cross Abstract: In an era defined by escalating climate change and the pervasive deployment of edge intelligence, the environmental cost of semiconductor manufacturin

Effect of Abstractions and Prompting Strategies on LLM-Guided High-Performance Optimizations

Model ReleasesDGX agent

arXiv:2608.08085v1 Announce Type: cross Abstract: Code performance optimization is a vital aspect of modern software development, as it enables faster response times and reduced resource usage. These

Efficient Cross-View Localization in 6G Space-Air-Ground Integrated Network

Local AiDGX agent

arXiv:2603.11398v2 Announce Type: replace-cross Abstract: Recently, visual localization has become an important supplement to improve localization reliability, and cross-view approaches can greatly en

EgoBrain: Synergizing Minds and Eyes For Human Action Understanding

ResearchDGX agent

arXiv:2506.01353v3 Announce Type: replace Abstract: The integration of brain-computer interfaces (BCIs), in particular electroencephalography (EEG), with artificial intelligence (AI) has shown tremend

EHR-MPC: Inference-Time Control for Sepsis Treatment with Generative Patient Digital Twins

SafetyDGX agent

arXiv:2607.08793v4 Announce Type: replace-cross Abstract: Sepsis is a leading cause of mortality, yet optimal treatment policies remain contested. Existing reinforcement learning (RL) approaches learn

El Agente Grafico: A Semantic Execution Runtime for Scientific Agents

AgentsDGX agent

arXiv:2602.17902v2 Announce Type: replace Abstract: Large language models (LLMs) can plan scientific workflows and generate code, but these capabilities do not specify how scientific state is validate

ElasticBack: Stealthy Conditional Backdoor in LLM-Agent Skills via Coupled Trigger-Rule Optimization

AgentsDGX agent

arXiv:2608.09577v1 Announce Type: new Abstract: Agent skills, bundles of instructions and resources that an LLM agent loads on demand, form an emerging supply chain where a single poisoned skill can p

ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language Models

Model ReleasesDGX agent

arXiv:2608.09548v1 Announce Type: cross Abstract: Large language models are increasingly deployed in education as tutors, teaching assistants, and content generators. These roles place demands that or

Embedding Trust: Semantic Isotropy Predicts Nonfactuality in Long-Form Text Generation

ApplicationsDGX agent

arXiv:2510.21891v2 Announce Type: replace-cross Abstract: To deploy large language models (LLMs) in high-stakes application domains that require substantively accurate responses to open-ended prompts,

EMMR: Emotion-Mediated Multimodal Reasoning for Personality Assessment in Asynchronous Video Interviews

ResearchDGX agent

arXiv:2608.07512v1 Announce Type: cross Abstract: Asynchronous Video Interviews (AVIs) have become increasingly popular for personality assessment. Recent large language models (LLMs) have shown poten

EmoPatient: An Emotion-Directed Patient Simulator for Realistic Palliative Care Communication Training

AgentsDGX agent

arXiv:2608.07495v1 Announce Type: cross Abstract: Effective communication during palliative care discussions is a critical clinical skill, yet training clinicians to manage complex patient emotions re

Emotion in an active inference model of human driving

ResearchDGX agent

arXiv:2608.07480v1 Announce Type: new Abstract: Active inference has emerged as a principled framework for modeling adaptive behavior by balancing goal-directed action with uncertainty reduction. It h

Emotion2Skill: Model-Internal Emotion Signals for Adaptive Skill Selection and Evolution

AgentsDGX agent

arXiv:2608.09248v1 Announce Type: new Abstract: Skill-based LLM agents select reusable procedures from an external library to solve complex tasks, yet their routing decisions rely entirely on text-lev

Energy-Structured Latent World Models with Neural Time Fields for Physically Constistent Open-World Motion Planning

SafetyDGX agent

arXiv:2608.09876v1 Announce Type: cross Abstract: Physically consistent motion planning remains a fundamental challenge in embodied AI, as generated trajectories must strictly conform to real-world ex

EnergyBridge: Benchmarking Household Energy Management, User Participation, and Grid Flexibility

Model ReleasesDGX agent

arXiv:2608.08691v1 Announce Type: new Abstract: Residential virtual power plants (VPPs) can provide grid flexibility by shifting household demand, but physical flexibility becomes dependable capacity

Enhanced Real-Time 6-DOF Extended Reality Catheter Tracking for Evaluating Potential Improvement in Efficiency, Precision, and Depth Perception for Cardiac Interventions

ResearchDGX agent

arXiv:2608.07606v1 Announce Type: cross Abstract: Despite advances in 3D ultrasound, most percutaneous cardiac interventions still rely on 2D visualization, limiting depth perception and spatial under

Enhancing Knowledge Tracing through Leakage-Free and Recency-Aware Embeddings

ResearchDGX agent

arXiv:2508.17092v2 Announce Type: replace-cross Abstract: Knowledge Tracing (KT) aims to predict a student's future performance based on their sequence of interactions with learning content. Many KT m

Enhancing Scientific Named Entity Recognition via Large Language Models: A Type-driven Multi-task Learning Approach

ResearchDGX agent

arXiv:2608.08636v1 Announce Type: cross Abstract: Scientific named entity recognition (SciNER) plays a crucial role in information extraction and knowledge discovery from scientific texts. Recently, l

Entropy-based Code Adversarial Translation for Real-world Repository Migration

Model ReleasesDGX agent

arXiv:2608.09273v1 Announce Type: new Abstract: LLMs have demonstrated strong capabilities in code generation and automated program repair, but migrating an entire repository rarely produces a runnabl

Epistemic Transfer in AI-Assisted Verification: A Framework and Evaluation Protocol

ResearchDGX agent

arXiv:2608.08882v1 Announce Type: cross Abstract: AI tools that help people judge online claims are usually evaluated while the tool is present. This paper asks a different question: after using such

Estimating Uncertainty in Galaxy Morphology Classification

ResearchDGX agent

arXiv:2608.08398v1 Announce Type: new Abstract: Astronomers classify galaxy morphology to investigate cosmic evolution. While deep foundation models are increasingly utilized in Galaxy Morphology Clas

Ethical Framework for Responsible Foundational Models in Medical Imaging

SafetyDGX agent

arXiv:2406.11868v2 Announce Type: replace-cross Abstract: The emergence of foundational models represents a paradigm shift in medical imaging, offering extraordinary capabilities in disease detection,

Evaluating Generative Time-Series Models on Data with Point Masses

Model ReleasesDGX agent

arXiv:2608.09692v1 Announce Type: cross Abstract: Many of the series that generative time-series models are benchmarked on place a large probability mass on a single value --- it does not rain, no rid

Evaluation of Motivational Interviewing Counsellors with Task-Aware Multi-Stage LLM-Based Simulated Clients

SafetyDGX agent

arXiv:2608.07499v1 Announce Type: cross Abstract: The development and benchmarking of Large Language Model (LLM)-based Motivational Interviewing (MI) counsellors now often rely on LLM-based simulated

Evidence-Grounded Forensic Reasoning for Detecting and Grounding Multi-Modal Media Manipulation

ResearchDGX agent

arXiv:2608.08009v1 Announce Type: cross Abstract: Fake news increasingly relies on cross-modal image-text forgeries, making transparent and verifiable reasoning chains an urgent need for Detecting and

Evidence-RL: Towards Evidence-intensive Visual Reasoning

Local AiDGX agent

arXiv:2608.08021v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) should answer from concrete image evidence rather than language priors, dataset shortcuts, or irrelevant visual context.

Evolving Safety Landscape of Multi-modal Large Language Models: A Survey of Emerging Threats and Safeguards

SafetyDGX agent

arXiv:2608.07535v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) integrate heterogeneous modalities through modality alignment and fusion, enabling stronger understanding an

Exact Network Surgery: Functional Invariance and Gradient Plasticity in Reactive Computational Graphs

ResearchDGX agent

arXiv:2607.16568v2 Announce Type: replace Abstract: Function-preserving network growth techniques such as Net2Net and progressive stacking expand a model's capacity without destroying its learned func

Exact Zarankiewicz Values On Two Finite Frontier Slices

ResearchDGX agent

arXiv:2608.08154v1 Announce Type: cross Abstract: The Zarankiewicz number Z(m,n,s,t) is the maximum number of edges in a bipartite graph with parts of orders m and n containing no copy of Ks,t. We giv

Experience-Sensitive Game Learning: A Behavioral Study of Humans and Language Agents

TutorialsDGX agent

arXiv:2608.07490v1 Announce Type: cross Abstract: Large language model agents are increasingly evaluated through games, but most benchmarks emphasize final outcomes rather than how players learn from

Expert-Guided Multimodal Fusion for Unified Emotion and Sentiment Analysis

Model ReleasesDGX agent

arXiv:2601.07565v2 Announce Type: replace-cross Abstract: Multimodal emotion understanding requires the integration of heterogeneous data sources, including text, audio, and visual modalities, while s

Explainable Machine Learning-Based Security and Privacy Protection Framework for Internet of Medical Things Systems

ApplicationsDGX agent

arXiv:2403.09752v4 Announce Type: replace-cross Abstract: The Internet of Medical Things transcends traditional medical boundaries, enabling a transition from reactive treatment to proactive preventio

Explaining, Verifying, and Aligning Semantic Hierarchies in Vision-Language Model Embeddings

SafetyDGX agent

arXiv:2603.26798v2 Announce Type: replace-cross Abstract: Vision-language model (VLM) encoders such as CLIP enable strong retrieval and zero-shot classification in a shared image-text embedding space,

Explore, Map, Remember, Decide: Are Embodied VLMs Ready for Safety-Critical Scenarios?

SafetyDGX agent

arXiv:2608.08077v1 Announce Type: new Abstract: Theory of Space framework (ToS) assesses the spatial understanding of curiosity-driven Vision-Language Models (VLMs) under partial observability. As AI

Exploring LLM Capabilities for Situational Understanding and COLREG compliance on real-world maritime navigation scenarios

ApplicationsDGX agent

arXiv:2608.08281v1 Announce Type: new Abstract: Recently, Large Language Models (LLMs) have shown considerable capability for situational understanding, reasoning, and decision making in different dom

exttt{DisMorph}: learning to disentangle technical distortions from true biological change

ResearchDGX agent

arXiv:2608.08173v1 Announce Type: cross Abstract: Longitudinal MRI enables sensitive measurement of structural brain change for studying aging and neurodegenerative disease. Deformable image registrat

FailForge: Distilling Procedural Competence from Persistent Failures into Code Agents

AgentsDGX agent

arXiv:2608.08570v1 Announce Type: new Abstract: Rejection sampling fine-tuning (RFT) is widely used to train code agents by generating trajectories on verifiable software engineering tasks, retaining

Fair on the Surface? Benchmarking Hidden-Output Fairness Gaps in LLM Recommenders

Model ReleasesDGX agent

arXiv:2608.08284v1 Announce Type: new Abstract: Fairness audits for LLM-based recommenders have largely focused on observable outputs, implicitly assuming that stable recommendations reflect stable in

FedA2L: Adaptive layer-wise learning rate adjustment in decentralized federated learning

Model ReleasesDGX agent

arXiv:2608.09208v1 Announce Type: cross Abstract: Decentralized intelligence systems with heterogeneous devices and limited coordination increasingly rely on decentralized federated learning (DFL). Ho

FedTVD: Balancing Data Quality and Quantity for Robust Federated Learning

Local AiDGX agent

arXiv:2608.09221v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training across distributed client devices while preserving data privacy. However, FL faces signif

FeedbackTrack: Visual-Cortex-Inspired Cross-Frame Feedback for Transformer Tracking

ResearchDGX agent

arXiv:2608.09369v1 Announce Type: cross Abstract: Visual object tracking requires effective temporal integration, yet most Transformer trackers still rely on predominantly feed-forward feature extract

FemWear: A Specialized Wearable Foundation Model for Women's Health

Model ReleasesDGX agent

arXiv:2608.08244v1 Announce Type: new Abstract: General wearable foundation models are pretrained across broad sensor streams and populations, but are not designed around women's-health tasks. We intr

Findings of the First Teaching Monster Challenge: A Benchmark of Pedagogical Content Knowledge in AI Agents

Model ReleasesDGX agent

arXiv:2608.08852v1 Announce Type: new Abstract: AI agents can now solve problems, answer like subject experts, and generate long-form multimodal content. However, whether they can adapt a lesson to fi

FitAQA: A Benchmark of Fitness Action Quality Assessment for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2608.08736v1 Announce Type: new Abstract: Fitness Action Quality Assessment (AQA) is important for intelligent sports training, yet the capabilities of Multimodal Large Language Models (MLLMs) i

Flow-by-Flow:Content-Judgment Bypass for Governing AI Output in High-Loss Domains

Model ReleasesDGX agent

arXiv:2608.07474v1 Announce Type: new Abstract: Prior work showed that human-in-the-loop oversight becomes structurally untenable in high-loss domains when AI output velocity V exceeds human cognitive

FoMoH: A clinically meaningful foundation model evaluation for structured electronic health records

Model ReleasesDGX agent

arXiv:2505.16941v4 Announce Type: replace-cross Abstract: Foundation models (FMs) promise to address core limitations of traditional supervised machine learning: (i) reliance on large amounts of label

ForestBench: A Unified Graph Framework for Evaluating Multi-Agent Collaboration

Model ReleasesDGX agent

arXiv:2608.08605v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on Large Language Models (LLMs) are proliferating rapidly, but their heterogeneous execution traces provide no common ba

Forgotten History or Test-of-Time? Retrospect and Prospect on RAG from an IR Perspective

SafetyDGX agent

arXiv:2608.08445v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is widely regarded as a novel paradigm born from the limitations of large language models (LLMs)--a mechanism to gr

Fourier Self-Supervision for Fine-Grained Generalized Category Discovery

ResearchDGX agent

arXiv:2608.08963v1 Announce Type: cross Abstract: Generalized Category Discovery aims to recognize known categories while identifying novel ones within unlabeled data. Existing methods, typically base

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering

ResearchDGX agent

arXiv:2608.08307v1 Announce Type: cross Abstract: Medical Visual Question Answering (VQA) requires aligning subtle visual evidence, including lesion texture, boundary sharpness, and diffuse density ch

FreSH: Frequency-Segmented Hierarchical Multi-Expert Framework for Multivariate Time Series Classification

Model ReleasesDGX agent

arXiv:2608.08207v1 Announce Type: cross Abstract: Multivariate Time Series Classification (MTSC) demands models that can effectively capture complex temporal patterns across multiple scales while rema

From Alignment to Synthesis Contrastive Volumetric Grounding for Text-to-CT Generation

SafetyDGX agent

arXiv:2506.00633v3 Announce Type: replace-cross Abstract: Generating semantically controllable 3D CT volumes from radiology reports requires more than a rich text encoder, it requires vision-language

From Evaluated Models to Evaluation Aids: A Multi-Evidence Study of LLM-Based Difficulty Calibration for Programming Examinations

Model ReleasesDGX agent

arXiv:2608.07523v1 Announce Type: cross Abstract: Difficulty differences across parallel-class programming examinations affect the fairness of course assessment. This study repositions large language

From Inaudible Inputs to Model Failures: Low-Frequency Safety Risks in LALMs

SafetyDGX agent

arXiv:2608.09158v1 Announce Type: cross Abstract: Large audio-language models (LALMs) have demonstrated strong capabilities in understanding diverse audio inputs. This diversity includes low-frequency

← Previous
1…678910…350
Next →