AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
7 Jul 2026

Estimating Individual Tree Height and Species from UAV Imagery

Model ReleasesDGX agent

arXiv:2603.23669v2 Announce Type: replace-cross Abstract: Accurate estimation of forest biomass, a major carbon sink, relies heavily on tree-level traits such as height and species. Unoccupied Aerial

evalci: A Python Library for Statistically Rigorous Comparison of Language Model Evaluations

AgentsDGX agent

arXiv:2607.04429v1 Announce Type: cross Abstract: The dominant practice in language model evaluation is to report a single accuracy number per model and declare the higher one better, without testing

Evaluating and Understanding Model Editing for Medical Vision Language Models

Model ReleasesDGX agent

arXiv:2607.05310v1 Announce Type: new Abstract: Model editing promises a fast, targeted way to correct post-deployment mistakes in medical vision-language models (VLMs) without costly retraining. Howe


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Evaluating Generative Agents with Actions Grounded in Socially Distributed Task Environments using Incognita

AgentsDGX agent

arXiv:2607.02975v1 Announce Type: new Abstract: Effective agency in social environments depends on when an agent seeks knowledge, when it acts, and whether its actions are justified by acquired inform

Evaluating LLM-Based Regression Test Generation

ResearchDGX agent

arXiv:2501.11086v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown tremendous promise in automated software engineering. In this paper, we investigate LLMs for just-in-t

Evaluating LLM Uncertainty in Long-Form Generation Using Deterministic Ground Truth

Model ReleasesDGX agent

arXiv:2607.03870v1 Announce Type: new Abstract: As LLMs generate increasingly long outputs, effective uncertainty estimation must identify errors at fine-grained levels rather than discard entire resp

Evaluating the Effect of Linguistic Relatedness on Cross-Lingual Transfer in Large Multilingual Automatic Speech Recognition

ResearchDGX agent

arXiv:2607.04814v1 Announce Type: cross Abstract: Extending automatic speech recognition (ASR) to low-resource African languages is constrained by the prohibitive demands of data collection at scale.

EventCoT: Event-centric Video Chain-of-thought for Reasoning Temporal Localization

Model ReleasesDGX agent

arXiv:2607.04872v1 Announce Type: cross Abstract: Reasoning temporal localization (RTL) requires a model to generate an answer that itself contains the time interval supporting it, so high-level reaso

EvoAgentBench: Benchmarking Agent Self-Evolution via Ability Transfer

Model ReleasesDGX agent

arXiv:2607.05202v1 Announce Type: new Abstract: Agent self-evolution in long-horizon LLM systems is largely procedural: useful experience is not merely stored information, but reusable procedures for

Evolutionary Guided Decoding: Iterative Value Refinement for LLMs

SafetyDGX agent

arXiv:2503.02368v4 Announce Type: replace-cross Abstract: While guided decoding, especially value-guided methods, has emerged as a cost-effective alternative for controlling language model outputs wit

EvoXplain: When Machine Learning Models Agree on Predictions but Disagree on Why -- Measuring Mechanistic Multiplicity Across Training Runs

ResearchDGX agent

arXiv:2512.22240v5 Announce Type: replace-cross Abstract: Machine learning models are primarily judged by predictive performance, especially in applied genomics, where explanations are read as biologi

Explainable AI for Screening Abuse-Related Trauma in Bangladeshi Children: A Training-Free Multimodal Framework Evaluated on Noise-Aware Synthetic Data

Model ReleasesDGX agent

arXiv:2607.04010v1 Announce Type: new Abstract: Bangladesh has an estimated 1.17 mental-health professionals per 100,000 population and only six child psychiatrists nationwide. No Bengali-language, cu

Explainable Bayesian deep learning through input-skip Latent Binary Bayesian Neural Networks

TutorialsDGX agent

arXiv:2503.10496v2 Announce Type: replace-cross Abstract: Modeling natural phenomena with artificial neural networks (ANNs) often provides highly accurate predictions. However, ANNs often suffer from

Explainable Novel Category Discovery in Semantic Concept Space

ResearchDGX agent

arXiv:2607.04548v1 Announce Type: cross Abstract: Novel category discovery aims to identify unseen classes from unlabeled data by transferring knowledge from labeled categories, but most existing meth

Explainable Reinforcement Learning for Adaptive Traffic Signal Control

SafetyDGX agent

arXiv:2607.03703v1 Announce Type: new Abstract: Reinforcement Learning (RL) has emerged as a powerful paradigm for adaptive traffic signal control. However, in safety-critical infrastructure like traf

Exploring Context-aware and LLM-driven Locomotion for Immersive Virtual Reality

ResearchDGX agent

arXiv:2504.17331v3 Announce Type: replace-cross Abstract: Locomotion plays a crucial role in shaping the user experience within virtual reality environments. In particular, hands-free locomotion offer

Exploring the Rashomon Set for Concept-Based Models

ResearchDGX agent

arXiv:2511.19636v2 Announce Type: replace-cross Abstract: In many machine learning problems, there may exist multiple models that achieve nearly identical predictive performance while relying on funda

Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tokens

SafetyDGX agent

arXiv:2508.04928v5 Announce Type: replace-cross Abstract: We propose a method to extend foundational monocular depth estimators (FMDEs), trained on perspective images, to fisheye images. Despite being

EyeMulator: Improving Code Language Models by Mimicking Human Visual Attention

TutorialsDGX agent

arXiv:2508.16771v3 Announce Type: replace-cross Abstract: Code Language Models (CodeLLMs) learn token importance from data correlations, whereas human developers attend selectively to semantically sal

Failures and Successes to Learn a Core Conceptual Distinction from the Statistics of Language

Model ReleasesDGX agent

arXiv:2607.04523v1 Announce Type: cross Abstract: Generic statements like 'tigers are striped' and 'cars have radios' communicate information that is, in general, true. However, while the first statem

FedACT: Federated Adaptive Coordinate Trust Modulation for Robust Transformer Training under Data Heterogeneity

Model ReleasesDGX agent

arXiv:2607.03763v1 Announce Type: cross Abstract: Federated Transformer training increasingly relies on local AdamW, whose adaptive updates can provide much stronger local progress than SGD-based trai

FedAvg for HAR: Exploring the Tradeoff Between Personalized and Generalization Accuracy

Local AiDGX agent

arXiv:2607.03334v1 Announce Type: cross Abstract: The federated learning (FL) paradigm fosters distributed pervasive computing combined with artificial intelligence techniques, allowing for optimized

Federated Learning for Object Detection: Enabling Collaborative Drone Learning Without Centralizing Data

Local AiDGX agent

arXiv:2607.02636v1 Announce Type: cross Abstract: Object detection is a fundamental capability for AI-driven perception in safety-critical drone and edge-vision systems, including disaster response, o

FedSPM: Routing-Enabled Federated Learning under Dual Heterogeneity via Semiparametric Mixture

Local AiDGX agent

arXiv:2607.04085v1 Announce Type: cross Abstract: Routing-prediction federated learning has emerged as a new paradigm that reframes inter-client heterogeneity as a resource for system-level intelligen

Fine-Grained Computation Offload for Off-the-Shelf Servers in Tens of Lines

ApplicationsDGX agent

arXiv:2607.02630v1 Announce Type: cross Abstract: Hardware accelerators now sit on the critical path of online serving. GPUs, FPGAs, and increasingly remote services such as hardware security modules,

Finite Reliability Representations: Noise-Calibrated Belief-Space Covers for Reliable Decision-Making

SafetyDGX agent

arXiv:2607.04019v1 Announce Type: cross Abstract: Physical sensing and actuation noise floors should inform how much belief resolution a decision-making system can reliably use. We introduce Finite Re

Fixed-Confidence Best-Arm Identification for Causal Mediation Analysis

ApplicationsDGX agent

arXiv:2607.04315v1 Announce Type: cross Abstract: This paper studies the problem of identifying the treatment that maximizes the expected natural direct potential outcome (NDPO), which captures the po

FlashBlock: Attention Caching for Efficient Long-Context Block Diffusion

ResearchDGX agent

arXiv:2602.05305v3 Announce Type: replace-cross Abstract: Generating long-form content, such as minute-long videos and extended texts, is increasingly important for modern generative models. Block dif

Flow-A11y: Flow-Aware Accessibility Testing

AgentsDGX agent

arXiv:2607.03100v1 Announce Type: cross Abstract: Modern web applications increasingly expose accessibility barriers through interaction flows rather than static page snapshots. Keyboard traps, focus

FM-ChangeNet: Learning Change through Pathwise Feature Transport

SafetyDGX agent

arXiv:2607.04750v1 Announce Type: new Abstract: We present FM-ChangeNet, a pathwise-supervised framework for change detection that reformulates bi-temporal reasoning as continuous transport in feature

Folding, Reasoning, and Scaling with Open-source Drug Discovery Engine

ResearchDGX agent

arXiv:2607.03787v1 Announce Type: new Abstract: Accurately modeling biomolecular interactions is a central bottleneck in biology and therapeutic discovery. Here, we introduce Open Drug Discovery Engin

Forethought: Verifiable Reasoning from Neurosymbolic Primitive Programming

Model ReleasesDGX agent

arXiv:2607.04096v1 Announce Type: new Abstract: Current agentic workflows usually involve decomposing user requests into sequences of tool calls with correctly resolved parameters, the results of whic

FORGE: Research-Trajectory Hijacking Attacks on Deep Research Agents

AgentsDGX agent

arXiv:2607.04718v1 Announce Type: new Abstract: Deep research agents decompose open-ended queries into subtasks, retrieve web evidence over multiple rounds, and synthesize long-form reports. This work

Formal Disco: Scalable Open-Ended Generation of Formally Verified Programs

Model ReleasesDGX agent

arXiv:2607.04631v1 Announce Type: new Abstract: The cost of producing code is rapidly diminishing with increasingly capable AI agents, while quality assurance of generated programs has not kept pace.

Foundations of Equivariant Deep Learning: Unifying Graph and Sheaf Neural Networks

ResearchDGX agent

arXiv:2607.03798v1 Announce Type: cross Abstract: Symmetry is everywhere in nature and society. Geometric deep learning exploits symmetries in data to improve the performance and efficiency of deep le

Framework of Thoughts: A Foundation Framework for Dynamic and Optimized Reasoning based on Chains, Trees, and Graphs

ResearchDGX agent

arXiv:2602.16512v2 Announce Type: replace Abstract: Prompting schemes such as Chain of Thought, Tree of Thoughts, and Graph of Thoughts can significantly enhance the reasoning capabilities of large la

From Arithmetic to Logic: The Resilience of Logic and Lookup-Based Neural Networks Under Parameter Bit-Flips

Model ReleasesDGX agent

arXiv:2603.22770v2 Announce Type: replace-cross Abstract: The deployment of deep neural networks (DNNs) in safety-critical edge environments necessitates robustness against hardware-induced bit-flip e

From Fixed to Free Cameras: Calibration-Free View-Robust Vision-Language-Action Model

SafetyDGX agent

arXiv:2607.05396v1 Announce Type: cross Abstract: Real-world robot deployment rarely maintains the training-stage camera setup, where cameras often experience repositioning or remounting depending on

From Judgments to Issues: Structured Extraction of Legal Reasoning with Citation-Hallucination Control

Model ReleasesDGX agent

arXiv:2607.03325v1 Announce Type: cross Abstract: We present an automated pipeline that decomposes Italian tax-court judgments into individual legal issues and extracts, for each issue, a structured X

From Mobile Data to Business Insights: An End-to-End Analytics Framework for Large-Scale Urban Mobility Analysis and Decision Support

ResearchDGX agent

arXiv:2607.03394v1 Announce Type: new Abstract: Real time location data derived from mobile applications is a powerful tool for addressing various urban challenges, including tourism planning, parking

From Raw Segmentations to Simulation-Ready Cardiac Meshes: An Automated Framework for Anatomical Reconstruction and Virtual Cohort Generation

Model ReleasesDGX agent

arXiv:2607.02564v1 Announce Type: cross Abstract: Computational models of the human heart are widely used to study electromechanical and fluid-dynamical cardiac function and to support applications su

From Regulation to Requirements: An Automated Requirement Derivation and Explanation Pipeline

Model ReleasesDGX agent

arXiv:2607.04448v1 Announce Type: cross Abstract: Ensuring software compliance with regulations such as the General Data Protection Regulation (GDPR) and the Artificial Intelligence Act (EU AI Act) po

From Tensor Buffer to Distributed Memory Hierarchy: A Survey of KV Cache Management for LLM Serving

ResearchDGX agent

arXiv:2607.02574v1 Announce Type: cross Abstract: The key-value (KV) cache has become a first-order memory object in LLM serving rather than a temporary per-request tensor. This survey classifies more

Full Glyph Images Beat Token Embeddings: A Controlled Study for Transformers

ResearchDGX agent

arXiv:2607.03994v1 Announce Type: cross Abstract: Modern language models generally represent text as sequences of discrete token embeddings, an assumption deeply rooted in current practice but rarely

Full-Stack FP4: Stable LLM Pretraining with Quantized Projections, Optimizers, and Attention

SafetyDGX agent

arXiv:2607.04422v1 Announce Type: cross Abstract: Recent NVFP4 pretraining methods mainly target transformer linear layers, leaving optimizer states, optimizer arithmetic and attention underexplored i

Fun-TSG: A Function-Driven Multivariate Time Series Generator with Variable-Level Anomaly Labeling

Model ReleasesDGX agent

arXiv:2604.14221v2 Announce Type: replace Abstract: Reliable evaluation of anomaly detection methods in multivariate time series remains an open challenge, largely due to the limitations of existing b

FuseMamba-VD: Dual Branch VideoMamba with Gated Class Token Fusion for Violence Detection

Model ReleasesDGX agent

arXiv:2506.03162v3 Announce Type: replace-cross Abstract: The rapid proliferation of surveillance cameras has increased the demand for automated violence detection. While CNNs and Transformers have sh

Fusion: A Framework for Unified Sequential Token AdaptatIon in VisiOn TraNsformers

ResearchDGX agent

arXiv:2607.02612v1 Announce Type: cross Abstract: Vision Transformers achieve strong image classification accuracy but process all image regions with nearly the same computation, even when many region

G2VD: Generalizable AI-Generated Video Detection via Counterfactual Intervention and Causal Disentanglement

SafetyDGX agent

arXiv:2607.04607v1 Announce Type: cross Abstract: The rapid advancement of AI-generated videos poses increasing security risks and calls for robust detectors with strong cross-domain generalization. A

GaP: A Graph-as-Policy Multi-Agent Self-Learning Harness For Variational Automation Tasks

SafetyDGX agent

arXiv:2607.05369v1 Announce Type: cross Abstract: For robots to work reliably in commercial and industrial applications, can recent advances in agentic coding systems combine interpretable robot progr

Gemma 4 Technical Report

Model ReleasesDGX agent

arXiv:2607.02770v1 Announce Type: cross Abstract: We introduce Gemma 4, a new generation of open-weight, natively multimodal language models in the Gemma model family. Designed to advance compute effi

Generative Semantic Multi-Object Tracking: A Large-Scale Benchmark and an MLLM-Driven Reasoning Framework

Model ReleasesDGX agent

arXiv:2601.06550v3 Announce Type: replace-cross Abstract: Semantic Multi-Object Tracking (SMOT) is evolving from purely geometric localization toward comprehensive video understanding. However, existi

Generative wave propagator

HardwareDGX agent

arXiv:2607.04440v1 Announce Type: cross Abstract: Seismic wavefield simulation is fundamental to seismology, but conventional finite-difference (FD) methods remain limited by numerical dispersion and

GenShin: Guiding Rational Liposome Design by Ranking Liposomal Protein Corona through a Docking-Pose-Free GNN

Model ReleasesDGX agent

arXiv:2504.13853v2 Announce Type: replace-cross Abstract: Rational design of lipid nanoparticles (LNPs) for tissue-specific delivery critically depends on predicting the composition of the protein cor

Geometry-Aware Motion Latents for Learning Robust Manipulation Policies

ApplicationsDGX agent

arXiv:2607.04714v1 Announce Type: cross Abstract: Learning motion latents for robotic manipulation heavily relies on extracting motion patterns from visual sequences, yet effective action abstractions

GeoSelect: Spatial-Program Execution for Training-Free Referring Remote Sensing Image Segmentation

Model ReleasesDGX agent

arXiv:2607.03869v1 Announce Type: cross Abstract: Referring remote sensing image segmentation isolates the object named by a natural-language expression in an aerial image. Existing training-free meth

GLM-5 Serving Parameter Tuning for OpenClaw: Single-Deployment MaaS Inference Optimization for Long-Context Agent Workloads

Model ReleasesDGX agent

arXiv:2607.02518v1 Announce Type: cross Abstract: OpenClaw requests are dominated by long, tool-augmented prefixes, including system prompts, conversation history, and tool outputs fed back into the c

Governed Individuation: Cryptographically Decoupling an Agent's Learning from Its Authority

Model ReleasesDGX agent

arXiv:2607.04613v1 Announce Type: new Abstract: Autonomous agents are moving from sandboxed text generators to operators of code, data, and physical infrastructure, and they increasingly learn while d

Governed MCP: Kernel-Level Tool Governance for AI Agents via Logit-Based Safety Primitives

Model ReleasesDGX agent

arXiv:2604.16870v2 Announce Type: replace-cross Abstract: AI agents increasingly call external tools (file system, network, APIs) through the Model Context Protocol (MCP). These tool calls are the age

Gradient Regularization Mitigates Reward Hacking in Reinforcement Learning from Human Feedback and Verifiable Rewards

SafetyDGX agent

arXiv:2602.18037v2 Announce Type: replace-cross Abstract: Reinforcement Learning from Human Feedback (RLHF) or Verifiable Rewards (RLVR) are two key steps in the post-training of modern Language Model

← Previous
1…8687888990…358
Next →