AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
30 Jun 2026

Dynamo: Dynamic Skill-Tool Evolution for Vision-Language Agents

Model ReleasesDGX agent

arXiv:2606.30185v1 Announce Type: new Abstract: Improving vision-language models (VLMs) on visual reasoning typically requires retraining or hand-designed prompts and tools. We present Dynamo, a train

Early Cue Precision Shapes Visual Shortcut Learning in Controlled Cue-Manipulation Benchmarks

Model ReleasesDGX agent

arXiv:2606.30344v1 Announce Type: cross Abstract: Visual classifiers can achieve high matched-distribution accuracy while relying on low-level cues that fail under conflict or suppression. We test whe

Early Warning Signals for OpenVLA Failure under Visual Distribution Shift

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.29699v1 Announce Type: cross Abstract: Vision Language Action models combine perception, language grounding, and control in a single policy, but their failures are hard to diagnose once vis

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks

Model ReleasesDGX agent

arXiv:2510.14207v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are powering a growing share of interactive web applications, yet remain vulnerable to misuse and harm. Prior jail

Edit in 2D, Verify in 3D: Reinforcement Learning for Multi-view Consistent Scene Editing

ApplicationsDGX agent

arXiv:2603.03143v2 Announce Type: replace-cross Abstract: Leveraging the priors of 2D diffusion models for 3D editing has emerged as a promising paradigm. However, multi-view consistency remains chall

Efficient RGB-T Object Detection via Sparse Cross-Modality Fusion

ResearchDGX agent

arXiv:2606.30215v1 Announce Type: cross Abstract: RGB-T detectors leverage the complementary strengths of visible and thermal infrared modalities, achieving robust performance under challenging condit

Efficient Spatio-Temporal Grounding with Multimodal Large Models via Second-Level Tracking and RL Verification

Local AiDGX agent

arXiv:2606.29023v1 Announce Type: cross Abstract: Spatio-temporal grounding in long videos requires precise temporal localization and robust object tracking conditioned on natural-language queries. Wh

EfficientUICoder: A Bidirectional Token Compression Framework for Efficient MLLM-Based UI Code Generation

ResearchDGX agent

arXiv:2509.12159v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models have demonstrated exceptional performance in UI2Code tasks, significantly enhancing website development effic

Em-ergence of the em-dash: a population-level rise in em-dash frequency in medRxiv preprints at the dawn of the large-language-model era

ResearchDGX agent

arXiv:2606.29540v1 Announce Type: cross Abstract: Large language models (LLMs) can leave subtle stylistic traces in assisted text; one of the most cited is the em-dash (Unicode U+2014). Yet no one has

Em-Garde: A Propose-Match Framework for Proactive Streaming Video Understanding

ResearchDGX agent

arXiv:2603.19054v2 Announce Type: replace-cross Abstract: Recent advances in Streaming Video Understanding has enabled a new interaction paradigm where models respond proactively to user queries. Curr

Emergence of Minimal Circuits for Indirect Object Identification in Attention-Only Transformers

Model ReleasesDGX agent

arXiv:2510.25013v2 Announce Type: replace-cross Abstract: Mechanistic interpretability aims to reverse-engineer large language models (LLMs) into human-understandable computational circuits. However,

EMPATH: A Multilingual Auditor-Judge Benchmark for Safety Evaluation of Emotional-Support Chatbots

Model ReleasesDGX agent

arXiv:2606.30256v1 Announce Type: new Abstract: Safety benchmarks often buy scalability by fixing the prompt, the language, and the turn structure. For emotional-support chatbots, that bargain hides p

ENC-ODE: Event-level Neurodegenerative Modeling in Continuous Time with Neural ODEs

ResearchDGX agent

arXiv:2606.30398v1 Announce Type: new Abstract: Accurately predicting the temporal evolution of clinical biomarkers is crucial for the early diagnosis and management of neurodegenerative diseases such

Enhanced Diffusion Sampling: Efficient Rare Event Sampling and Free Energy Calculation with Diffusion Models

HardwareDGX agent

arXiv:2602.16634v2 Announce Type: replace-cross Abstract: The rare-event sampling problem has long been the central limiting factor in molecular dynamics (MD), especially in biomolecular simulation. R

Ensemble Learning Based Classification Algorithm Recommendation

Model ReleasesDGX agent

arXiv:2101.05993v2 Announce Type: replace-cross Abstract: Selecting an appropriate classification algorithm for a given data set remains a challenging problem in data mining and machine learning. Exis

Entity Binding Failures in Tool-Augmented Agents

SafetyDGX agent

arXiv:2606.30531v1 Announce Type: new Abstract: Tool-augmented language-model agents are often evaluated by whether they select the correct tool, produce valid API arguments, and complete the requeste

Entropy-Gated Latent Recursion

ResearchDGX agent

arXiv:2606.16620v2 Announce Type: replace-cross Abstract: Inference-time scaling has become the dominant lever for improving language-model reasoning, but existing methods derive rollout diversity fro

Estimating Grammatical Gender Directions in Contextual Embeddings under Controlled and Natural Contexts

SafetyDGX agent

arXiv:2606.30152v1 Announce Type: cross Abstract: Contextual language models conflate grammatical gender and social semantic bias in gendered languages such as Spanish. Existing gender debiasing appro

EVAF: A Test-Retest Protocol for Selective Parametric Consolidation

Model ReleasesDGX agent

arXiv:2606.29916v1 Announce Type: cross Abstract: Long-running language agents need mechanisms for deciding which experiences should persist after the working context is gone. Retrieval systems can re

EvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures

Model ReleasesDGX agent

arXiv:2606.30219v1 Announce Type: new Abstract: LLM evaluation and AI safety face a shared measurement problem: benchmark scores, reward-model signals, and reported safety metrics can improve while th

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

Model ReleasesDGX agent

arXiv:2507.05257v4 Announce Type: replace-cross Abstract: Recent benchmarks for Large Language Model (LLM) agents primarily focus on evaluating reasoning, planning, and execution capabilities, while a

Event-Conditioned Diagnostics of Kinematic, Contact, and Object-Permanence Fields in Passive Object-State World Models

TutorialsDGX agent

arXiv:2606.28455v1 Announce Type: cross Abstract: World models can predict future physical states, but prediction accuracy alone does not explain how physical information is organized and used inside

Evidence-Based Text-Conditioned 3D CT Synthesis for Ovarian Cancer

SafetyDGX agent

arXiv:2606.28980v1 Announce Type: cross Abstract: Ovarian cancer is frequently diagnosed at an advanced stage, making preoperative contrast-enhanced computed tomography (CT) central to staging and sur

Evidence-Driven LLM Agent for C-to-Synthesizable-C Conversion and Verification

AgentsDGX agent

arXiv:2606.28409v1 Announce Type: cross Abstract: Software-compilable C programs routinely fail to complete the four-stage pipeline of a high-level synthesis (HLS) toolchain -- compilation, C simulati

Evidence-Informed LLM Beliefs for Continual Scientific Discovery

ResearchDGX agent

arXiv:2606.29182v1 Announce Type: new Abstract: Open-ended scientific discovery with large language models (LLMs) increasingly operates as a long-horizon loop of hypothesis search and verification, wh

Evolutional Math: Cross-Validated Island-Model Genetic Programming for Interpretable Symbolic Regression on Small, Wide Datasets

Model ReleasesDGX agent

arXiv:2606.28381v1 Announce Type: cross Abstract: Symbolic regression via genetic programming routinely fails on small, wide datasets - a regime common in clinical-trial monitoring, biostatistics, and

Exit-and-Join Dynamics and Equilibrium in Continuum Cooperative Games

AgentsDGX agent

arXiv:2606.28824v1 Announce Type: cross Abstract: This paper develops a continuum theory of exit-and-join coalition dynamics in nonatomic cooperative games. We extend the Aumann-Shapley value and the

Experience-Evolving Multi-Turn Tool-Use Agent with Hybrid Episodic-Procedural Memory

SafetyDGX agent

arXiv:2512.07287v3 Announce Type: replace-cross Abstract: As intents unfold and environments change, multi-turn agents face continuously shifting decision contexts. Although reusing past experience is

Experience Graphs: The Data Foundation for Self-Improving Agents

AgentsDGX agent

arXiv:2606.29823v1 Announce Type: cross Abstract: The database community has repeatedly advanced the state of the art by recognizing that new workloads demand new system architectures. We argue that l

Expert Evaluation of Clinical AI Tools on Real Point-of-Care Clinical Queries

Model ReleasesDGX agent

arXiv:2606.28960v1 Announce Type: new Abstract: Physicians now pose millions of clinical questions to AI tools each week, yet these tools are evaluated largely on hypothetical or exam-style questions,

Explaining Attention with Program Synthesis

Model ReleasesDGX agent

arXiv:2606.19317v2 Announce Type: replace-cross Abstract: A longstanding goal of research on interpretable deep learning is to replace opaque neural computations with human-meaningful symbolic descrip

Exploiting Local Flatness for Efficient Out-of-Distribution Detection

Model ReleasesDGX agent

arXiv:2606.29952v1 Announce Type: cross Abstract: Detecting out-of-distribution (OOD) data is crucial for reliable machine learning deployment. Among detection strategies, post-hoc methods are particu

Exploration and Online Transfer with Behavioral Foundation Models

SafetyDGX agent

arXiv:2606.29980v1 Announce Type: new Abstract: Zero-shot Transfer in Reinforcement Learning (RL) aims to train an agent that can generate optimal policies for any reward function, without additional

Exploring Motivations for Algorithm Mention in the Domain of Natural Language Processing: A Deep Learning Approach

ResearchDGX agent

arXiv:2606.29859v1 Announce Type: cross Abstract: With the rise of data-intensive science, algorithms have become central to scientific research. In academic papers, algorithms are mentioned for diffe

Exploring the Value of Diverse LLM Explanations in Introductory Programming

ApplicationsDGX agent

arXiv:2606.28882v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown the potential to generate code explanations that surpass those of peers in quality, offering promising opportu

EyeMVP: OCT-Informed Fundus Representation Learning via Paired CFP--OCT Pretraining

TutorialsDGX agent

arXiv:2606.15129v2 Announce Type: replace-cross Abstract: Color fundus photography (CFP) is the mainstay of large-scale retinal screening, but its diagnostic capacity is limited by the lack of depth-r

FacePlex: Full-Duplex Joint Speech-Facial Motion Generation for Conversational Avatars

ResearchDGX agent

arXiv:2606.30145v1 Announce Type: new Abstract: Natural face-to-face conversation requires real-time speech generation together with synchronized facial motion. Existing systems only partially address

FADE: Mitigating Hallucinations by Reducing Language-Prior Dominance in Large Vision-Language Models

ResearchDGX agent

arXiv:2606.29431v1 Announce Type: new Abstract: Despite the impressive capabilities of Large Vision-Language Models (LVLMs), they remain susceptible to hallucination, generating content inconsistent w

Fairness Attacks on Recommender Systems

SafetyDGX agent

arXiv:2606.29064v1 Announce Type: cross Abstract: The unfairness of recommender systems has become a topic of concern due to its significant social and ethical implications. Although existing works ha

FalconTrack: Photorealistic Auto-Labeled Perception and Physics-Aware Vision-Based Aerial Tracking

ResearchDGX agent

arXiv:2606.29783v1 Announce Type: cross Abstract: Vision-based aerial tracking is critical in GPS-denied environments. Reliable perception for tracking depends on large-scale labeled data, yet most ph

Fast and Accurate Outlier-Aware LiDAR Super-Resolution for SLAM Applications

ResearchDGX agent

arXiv:2606.28607v1 Announce Type: cross Abstract: This work tackles the challenge of enhancing low-resolution LiDAR sensors for SLAM applications through a novel Deep Unrolling-based Super-Resolution

Fast Enough to Act: Spatio-Temporal Visual Token Merging for Low-Latency Robotic VLMs and VLAs

SafetyDGX agent

arXiv:2606.29350v1 Announce Type: cross Abstract: Vision-language models and vision-language action models endow the robot with unprecedented capabilities. However, the input of video and high-resolut

Fast Wireless Foundation Models with Early-Exits

ResearchDGX agent

arXiv:2606.29640v1 Announce Type: cross Abstract: While wireless foundation models (FMs) are demonstrating strong potential to enable AI-Native 6G networks, their high computational cost remains a cri

Faults in Our Formal Benchmarking: Dataset Defects and Evaluation Failures in Lean Theorem Proving

TutorialsDGX agent

arXiv:2606.29493v1 Announce Type: new Abstract: Benchmarks for LLM-assisted theorem proving in Lean are often treated as intrinsically reliable because every solved instance comes with a machine-check

FD^2: A Dedicated Framework for Fine-Grained Dataset Distillation

ResearchDGX agent

arXiv:2603.25144v2 Announce Type: replace-cross Abstract: Dataset distillation (DD) compresses a large training set into a small synthetic set, reducing storage and training cost, and has shown strong

Feature-level Interaction Explanations in Multimodal Transformers

ResearchDGX agent

arXiv:2603.13326v2 Announce Type: replace-cross Abstract: Multimodal Transformers often produce predictions without clarifying how different modalities jointly support a decision. Most existing multim

Federated Learning with Energy-Based Structured Probabilistic Inference

ResearchDGX agent

arXiv:2606.30161v1 Announce Type: cross Abstract: Federated learning typically aggregates client updates using fixed or heuristic weighting rules, which can be suboptimal when clients have heterogeneo

FedLAS: Feature-Modulated Bidirectional Label Smoothing for Neural Network Calibration

ResearchDGX agent

arXiv:2606.28654v1 Announce Type: cross Abstract: Deep Neural Network (DNN) classifiers suffer from poor calibration when their softmax outputs (predictive confidence) deviate from the empirical likel

Few-class Fidelity: Evaluating Explanations of Real-conditions CNN classifiers with Optimized Perturbations

Local AiDGX agent

arXiv:2606.28391v1 Announce Type: cross Abstract: The wide use of Convolutional Neural Networks (CNN) in numerous domains and real-world classification applications is justified by their high precisio

Few-Shot Domain Incremental Learning via Continual Vision-Language Consolidation

Model ReleasesDGX agent

arXiv:2606.30190v1 Announce Type: cross Abstract: Existing domain-incremental learning (DIL) strategies call for massive amounts of data to adapt to new domains and suffer from the overfitting problem

FFAvatar: Feed-Forward 4D Head Avatar Reconstruction from Sparse Portrait Images

ResearchDGX agent

arXiv:2606.30347v1 Announce Type: cross Abstract: We present FFAvatar, a Transformer-based 3D Gaussian framework for fast construction of high-quality and animatable 4D head avatars from one or more r

Field Order Should Not Matter: Permutation-Invariant Embedding Model Fine-Tuning for Structured Metadata Retrieval

Model ReleasesDGX agent

arXiv:2606.30473v1 Announce Type: cross Abstract: We study retrieval over catalogs of structured metadata, where each record is a small schema whose fields answer different kinds of query. Embedding a

Financing Artificial Intelligence Infrastructure: Mapping AI Infrastructure Investment and Compute Governance Across Africa

ApplicationsDGX agent

arXiv:2606.28404v1 Announce Type: cross Abstract: Artificial intelligence depends on large-scale compute resources and their supporting infrastructure. However, AI governance debates treat compute pri

Fine-Tuning General-Purpose Large Language Models for Agricultural Applications:A Reproducible Framework and Evaluation Protocol Based on Qwen3-8B

Model ReleasesDGX agent

arXiv:2606.28992v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) have demonstrated strong abilities in opendomain question answering, information extraction, and text gen

First-Order Temporal Logic Tensor Networks

ResearchDGX agent

arXiv:2606.29972v1 Announce Type: new Abstract: Most of the existing neuro-symbolic AI methods focus on the scenario of static knowledge where objects do not change according to a temporal dimension.

Fisher-Routed Mixture of Experts for Federated Class-Incremental Learning

ResearchDGX agent

arXiv:2606.28835v1 Announce Type: cross Abstract: Federated Learning (FL) emerged as a promising distributed machine learning paradigm. However, extending FL to the class incremental learning scenario

FLAME 3 Dataset: Unleashing the Power of Radiometric Thermal UAV Imagery for Wildfire Management

ResearchDGX agent

arXiv:2412.02831v2 Announce Type: replace-cross Abstract: The increasing accessibility of radiometric thermal imaging sensors for unmanned aerial vehicles (UAVs) offers significant potential for advan

FlatLands: Generative Floormap Completion From a Single Egocentric View

Model ReleasesDGX agent

arXiv:2603.16016v2 Announce Type: replace-cross Abstract: A single egocentric image typically captures only a small portion of the floor, yet a complete metric traversability map of the surroundings w

Flow Matching in Feature Space for Stochastic World Modeling

Model ReleasesDGX agent

arXiv:2606.29059v1 Announce Type: cross Abstract: World modeling requires forecasting uncertain futures while preserving information useful for downstream perception. Existing visual world models ofte

Flow Reasoning Models: Scaling Reasoning Through Iterative Self-Refinement

ResearchDGX agent

arXiv:2606.29150v1 Announce Type: new Abstract: Discrete flow models have recently shown promising performance on few-step text generation; however, when naively applied to structured reasoning tasks

← Previous
1…110111112113114…358
Next →