AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
7 Jul 2026

Criterion-Conditional In-Context Learning: Evaluating Criterion-Shift Adaptation in Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.02575v1 Announce Type: cross Abstract: Vision-language models can perform new tasks without parameter updates through in-context learning (ICL), whose core mechanism is utilizing the suppor

CritiqueDriveVLM: From Verifier-Guided Reinforcement Learning to Latent Thought Distillation for Autonomous Driving

Model ReleasesDGX agent

arXiv:2607.04179v1 Announce Type: cross Abstract: End-to-end Vision-Language Models (VLMs) show immense potential in autonomous driving. However, standard Supervised Fine-Tuning (SFT) often suffers fr

CRODA-ST: Single-Target Cross-Receiver Open-Set Radio Fingerprint Recognition

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.02567v1 Announce Type: cross Abstract: Radio frequency fingerprint identification (RFFI) provides a physical-layer credential for Internet of Things devices, but open-set decisions become f

CRRL: A Causality-Based Reinforcement Learning Framework for Autonomous System Recovery

SafetyDGX agent

arXiv:2607.03177v1 Announce Type: cross Abstract: Traditional reinforcement learning (RL) for recovery in autonomous systems lacks causal understanding and generalizes poorly to novel failure scenario

CSB: A Counting and Sampling tool for Bit-vectors

ResearchDGX agent

arXiv:2607.04142v1 Announce Type: cross Abstract: Satisfiability modulo theory (SMT) solvers have significantly advanced automated reasoning due to their effectiveness in solving problems across vario

CuBAS: Information Geometric Curvature-Based Adaptive Sampling for Supervised Classification

Model ReleasesDGX agent

arXiv:2607.03145v1 Announce Type: cross Abstract: The informativeness of a training set is as consequential as its size, yet most sampling strategies remain agnostic to the intrinsic geometry of the d

Dashboard2Code: Evaluating Multimodal Models on Reconstructing Interactive Dashboards

Model ReleasesDGX agent

arXiv:2607.04727v1 Announce Type: cross Abstract: Automatic data visualization generation has advanced rapidly with multi-modal large language models, yet existing efforts largely focus on static char

Data Driven Optimization of GPU efficiency for Distributed LLM-Adapter Serving

HardwareDGX agent

arXiv:2602.24044v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) adapters enable low-cost model specialization, but introduce complex caching and scheduling challenges in distribut

Decentralised Federated Learning over Temporal Networks: The Role of Heterogeneities

Local AiDGX agent

arXiv:2607.03171v1 Announce Type: cross Abstract: Decentralised federated learning, based on peer-to-peer communication, is increasingly proposed for on-device training of machine learning models, pro

Decentralized Aggregation of LLM Predictions via Wagering Mechanisms

SafetyDGX agent

arXiv:2607.04389v1 Announce Type: new Abstract: It is increasingly common to aggregate predictions from multiple LLMs, each with domain expertise or access to private tools and data, to improve collec

DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences

Local AiDGX agent

arXiv:2607.02551v1 Announce Type: cross Abstract: Video multimodal large language models have made strong progress on open-ended video understanding, but they still lack precise local spatiotemporal p

Demonstrating Generalization Failures via Mixtures of Conditional Policies

SafetyDGX agent

arXiv:2607.03478v1 Announce Type: new Abstract: Post-training of frontier language models is conducted on curated task suites, and inevitably leaves a distribution shift between training and deploymen

Deriving Neural Scaling Laws from the statistics of natural language

Model ReleasesDGX agent

arXiv:2602.07488v3 Announce Type: replace-cross Abstract: Despite the fact that experimental neural scaling laws have substantially guided empirical progress in large-scale machine learning, no existi

DETECT-3B-Omni is Agnostic of Content and Demographics

ResearchDGX agent

arXiv:2607.03418v1 Announce Type: cross Abstract: A trustworthy and GDPR-compliant deepfake audio detector must base its decisions on acoustic artifacts, not on what is being said or who is speaking.

Detecting Answer-Driven Reasoning in LLM-Based Educational Tutors via Truncated Chain-of-Thought Auditing

TutorialsDGX agent

arXiv:2607.04572v1 Announce Type: new Abstract: Large language model (LLM) tutors often produce fluent step-by-step explanations, but a correct and pedagogically formatted response does not guarantee

Detecting Architectural Drift in Safety-Critical Firmware through Runtime Trace Analysis

SafetyDGX agent

arXiv:2607.03135v1 Announce Type: cross Abstract: Maintaining consistency between architectural design and runtime-observed behavior is challenging in long-lived safety-critical firmware. This paper p

Detecting Hallucinations in Retrieval-Augmented Generation through Grounding-Aware Sensitivity by Perturbation (GASP)

ResearchDGX agent

arXiv:2607.04223v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) reduces but does not eliminate hallucination, and existing detectors return a single answer-level score that does

Determinants and Limits of LLM Security-Tool Orchestration: A Study with HexStrike-AI

Model ReleasesDGX agent

arXiv:2607.02873v1 Announce Type: cross Abstract: Large language model agents driving security tool suites over the Model Context Protocol are increasingly common. Yet the factors that bound their cap

Developing an LLM-Based Feedback System Grounded in Evidence-Centered Design to Support Physics Problem Solving

ResearchDGX agent

arXiv:2512.10785v3 Announce Type: replace-cross Abstract: Generative AI offers new opportunities for individualized and adaptive learning, e.g., through large language model (LLM)-based feedback syste

Diagnosing Aerial-View Object Detectors with Foundational Image Generative Models

TutorialsDGX agent

arXiv:2607.02718v1 Announce Type: cross Abstract: Recent advances in large-scale image generative models enable photorealistic scene synthesis with controllable attributes. Beyond data augmentation, t

Differential Amplifier-Inspired AmpAttention for Multi-View Robotic Manipulation

ApplicationsDGX agent

arXiv:2607.02845v1 Announce Type: cross Abstract: Multi-view robotic manipulation methods with the attention mechanism have recently achieved significant progress in both training efficiency and task

Differentiate the Evaluator, Not the Program: An Efficient Runtime Representation for Neuro-Symbolic Learning

Model ReleasesDGX agent

arXiv:2607.03574v1 Announce Type: cross Abstract: AI systems increasingly propose executable scientific models whose value depends on both their symbolic structure and their fitted continuous paramete

Diffusion-Guided Uncertainty-Aware Delayed Policy Optimization

SafetyDGX agent

arXiv:2607.05064v1 Announce Type: new Abstract: Reinforcement learning in real world environments often suffers from severe performance degradation due to delayed feedback. Existing approaches typical

Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval

ResearchDGX agent

arXiv:2607.04605v1 Announce Type: cross Abstract: Multi-vector vision-language retrieval preserves fine-grained visual evidence through maximum-similarity late interaction, but dense image-side tokens

Do GUI Agents Believe Their Eyes? Diagnosing State-Belief Reliance on Pixels versus Structure

AgentsDGX agent

arXiv:2607.04334v1 Announce Type: new Abstract: Multimodal GUI agents read an interface through two redundant channels: the rendered pixels of a screenshot and a serialized structure such as a DOM or

Do Vision-Language-Action Models Mean What They Say? On the Role of Faithfulness in Embodied Reasoning

SafetyDGX agent

arXiv:2607.04681v1 Announce Type: cross Abstract: Embodied Chain-of-Thought has emerged as a promising mechanism to enhance robot decision-making and interpretability in black-box Vision-Language Acti

Domain Knowledge-Informed Self-Supervised Representations for Workout Form Assessment

TutorialsDGX agent

arXiv:2202.14019v3 Announce Type: replace-cross Abstract: Maintaining proper form while exercising is important for preventing injuries and maximizing muscle mass gains. Detecting errors in workout fo

Don't Blame the Large Language Model: How Scaffolding Evolution Shapes Coding Agent Quality

Model ReleasesDGX agent

arXiv:2607.03691v1 Announce Type: cross Abstract: Coding agents, autonomous systems that use large language models (LLMs) to resolve software engineering tasks, rely on agentic scaffolding: a middlewa

Don't Wait to Reply: Towards Responsive yet Thoughtful Dialogue through Proactive Thinking

ResearchDGX agent

arXiv:2607.03093v1 Announce Type: cross Abstract: Thinking has emerged as a critical capability for Large Language Models (LLMs) tackling complex tasks. However, its reactive nature, where reasoning i

dOPSD: On-Policy Self-Distillation for Diffusion Language Models

SafetyDGX agent

arXiv:2607.04428v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising a masked sequence, offering a parallel alternative to autoregressive mo

DOSE-I: A Multimodal Biosignal Dataset of Procedural Sedation for Endoscopy -- Technical Report

ResearchDGX agent

arXiv:2607.02570v1 Announce Type: cross Abstract: In this document, we describe characteristics and technical details of the multimodal biosignal dataset DOSE-I of procedural sedation for endoscopy pu

Double Fuzzy Probabilistic Interval Linguistic Term Set and a Dynamic Fuzzy Decision Making Model based on Markov Process with tts Application in Multiple Criteria Group Decision Making

ResearchDGX agent

arXiv:2111.15255v2 Announce Type: replace-cross Abstract: The probabilistic linguistic term has been proposed to deal with probability distributions in provided linguistic evaluations. However, becaus

Double-Helix Active Geometry: LiDAR-Anchored Multi-View Depth with Selective Abstention

SafetyDGX agent

arXiv:2607.02561v1 Announce Type: cross Abstract: Consumer depth sensors such as the LiDAR scanner on recent iPhones provide metric range, but their useful range is short and their returns are sparse.

DrugAgent: Reliable Multi-Agent Integration of Conflicting Biomedical Evidence for Drug-Target Interaction Assessment

Model ReleasesDGX agent

arXiv:2408.13378v5 Announce Type: replace Abstract: Workflows in drug-target interaction (DTI) assessment require integrating heterogeneous data from predictive models, curated resources, and observat

DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation

Model ReleasesDGX agent

arXiv:2607.05147v1 Announce Type: new Abstract: Speculative decoding accelerates Large Language Model (LLM) inference by decoupling draft generation from target verification. While recent parallel dra

DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation

SafetyDGX agent

arXiv:2607.04927v1 Announce Type: cross Abstract: World Action Models (WAMs) provide a promising alternative to Vision-Language-Action (VLA) policies by using video-based world modeling as dense super

DualView: Preventing Indirect Prompt Injection in Personal AI Agents

Model ReleasesDGX agent

arXiv:2607.03821v1 Announce Type: cross Abstract: Personal AI agents that run on the user's local machine, such as OpenClaw, automate daily tasks including web search, email, and file management. Thei

DynamixSFT: Dynamic Mixture Optimization of Instruction Tuning Collections

ResearchDGX agent

arXiv:2508.12116v2 Announce Type: replace-cross Abstract: As numerous instruction-tuning datasets continue to emerge, dynamically balancing and optimizing their mixtures has become a critical challeng

DynaVieW: Schema-Guided World Modeling for Understanding Hierarchical Visual Dynamics

ResearchDGX agent

arXiv:2607.04112v1 Announce Type: cross Abstract: Multimodal LLMs struggle to systematically model the temporal evolution of visual scenes in videos or multi-image sequences. Such inputs require model

Echoes of Unrest: A Multimodal NLP Framework for Early Warning of Fake News and Violence-Driven Mob Activity

Model ReleasesDGX agent

arXiv:2607.02734v1 Announce Type: cross Abstract: Rapid growth in social media has transformed global communication by enabling fast information exchange, but it has also accelerated the spread of mis

EEG-SpikeAgent: Agentic Closed-Loop Program Synthesis for Automated EEG Spike Detection

AgentsDGX agent

arXiv:2607.04558v1 Announce Type: cross Abstract: Automated detection of interictal epileptiform discharges in scalp electroencephalography (EEG) is clinically important, but recent high-performing de

Effectiveness of LLM-based Software Diversity for Reliability Improvement -- an Empirical Study

ResearchDGX agent

arXiv:2607.03174v1 Announce Type: cross Abstract: Software diversity has been extensively studied as a means of reducing the risk of common-mode failures. Classic work showed that the central issue is

Efficient bias mitigation in T2I diffusion models using Concept Graphs

SafetyDGX agent

arXiv:2607.03397v1 Announce Type: new Abstract: Text-to-Image diffusion models often propagate harmful bias inherited from the training data. Existing bias mitigation techniques typically intervene on

Efficient Decentralized Multi-task Dataset Valuation via Model Merging

Model ReleasesDGX agent

arXiv:2607.03346v1 Announce Type: cross Abstract: Accurate and efficient dataset valuation is essential for enabling fair and transparent data marketplaces, especially when multiple contributors provi

Efficient Discovery of Conditional Dependencies with Desbordante

ResearchDGX agent

arXiv:2607.04030v1 Announce Type: cross Abstract: Conditional functional dependencies (CFDs) are functional dependencies with a restricted scope: they specify the context in which a dependency holds a

Efficient Flow Matching for Sparse-View CT Reconstruction

ResearchDGX agent

arXiv:2603.00205v2 Announce Type: replace-cross Abstract: Generative models, particularly Diffusion Models (DM), have shown strong potential for Computed Tomography (CT) reconstruction serving as expr

Efficient Perception in Automotive Detection and Tracking Using Neuromorphic Computing

AgentsDGX agent

arXiv:2607.04921v1 Announce Type: cross Abstract: Deep learning algorithms are notorious for their high carbon footprint and computational demands that limit their deployment on edge devices and raise

EGRA:Toward Enhanced Behavior Graphs and Representation Alignment for Multimodal Recommendation

SafetyDGX agent

arXiv:2508.16170v2 Announce Type: replace-cross Abstract: MultiModal Recommendation (MMR) systems have emerged as a promising solution for improving recommendation quality by leveraging rich item-side

Elastic Gang: Per-Token Membership Change for a Hard-Barriered LLM Inference Gang Co-Scheduled with OS Processes

Local AiDGX agent

arXiv:2607.04668v1 Announce Type: cross Abstract: On-device LLM decoding is a hard-barriered CPU-SIMD computation that wants every core for milliseconds per token, while the rest of the OS wants those

ELBO-T2IAlign: A Generic ELBO-Based Method for Calibrating Pixel-level Text-Image Alignment in Diffusion Models

SafetyDGX agent

arXiv:2506.09740v2 Announce Type: replace-cross Abstract: Diffusion models excel at image generation. Recent studies have shown that these models not only generate high-quality images but also encode

ELiTeFormer: An Efficient Transformer for FPGAs

Model ReleasesDGX agent

arXiv:2607.03652v1 Announce Type: cross Abstract: Transformer blocks are prevalent in large language model (LLM) but present deployment challenges due to their challenging computational and memory dem

ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval

Model ReleasesDGX agent

arXiv:2606.20280v2 Announce Type: replace-cross Abstract: Leveraging Multimodal Large Language Models (MLLMs) via contrastive learning has become a mainstream paradigm for improving the performance of

Embodied Operators and Benchmarking: Toward Reusable and Deployable Embodied Intelligence Systems

Model ReleasesDGX agent

arXiv:2607.03283v1 Announce Type: new Abstract: Embodied intelligence systems require not only end-to-end policy models, but also reusable functional modules that transform multimodal observations, ro

EmCom-Diffusion: Probing Visual Reflection in Emergent Languages via Image Generation

ResearchDGX agent

arXiv:2607.03752v1 Announce Type: cross Abstract: Measuring the extent to which emergent languages encode the visual content of their inputs is an open problem. We refer to this property as visual ref

Empirical Computation: Prompting versus Programming

TutorialsDGX agent

arXiv:2503.10954v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) agents can solve *any* computational problem *without* an algorithm in a runtime *independent* of the computational

Enhanced Feature Extraction for IoT Network Intrusion Detection Using GNNs and KAN

ResearchDGX agent

arXiv:2607.02981v1 Announce Type: cross Abstract: Recent advancements in the Internet of Things (IoT) emphasize the urgent need for advanced network security, as IoT networks feature dynamic topologie

Enhancement of E-commerce Sponsored Search Relevancy with LLM

Model ReleasesDGX agent

arXiv:2607.03886v1 Announce Type: cross Abstract: Sponsored search plays a crucial role as a revenue stream for search engines, wherein advertisers competitively bid on keywords that align with the us

Enhancing Implicit Neural Representations with Image Feature Embedding for Unsupervised Cardiac Cine MRI Reconstruction

ResearchDGX agent

arXiv:2607.04069v1 Announce Type: cross Abstract: Cardiac cine Magnetic Resonance Imaging (MRI) is a critical diagnostic tool that provides dynamic insights for radiologists. To accelerate acquisition

Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning

SafetyDGX agent

arXiv:2506.05716v2 Announce Type: replace-cross Abstract: Deep Q-Networks (DQN) can suffer from overestimation bias because bootstrapped targets use a maximisation operation over noisy value estimates

EPRA U-Net: An Efficient Pyramid Residual Attention Framework for Accurate Infarct Segmentation in Diffusion-Weighted MRI

ResearchDGX agent

arXiv:2607.03568v1 Announce Type: cross Abstract: Objective: Accurate identification of acute ischemic infarcts on diffusion-weighted magnetic resonance imaging (DWI) is a critical prerequisite for re

← Previous
1…8586878889…358
Next →