AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
9 Jun 2026

EgoAERO: Learning Dexterous Manipulation from a Single Egocentric Video without Object Assets

SafetyDGX agent

arXiv:2606.08057v1 Announce Type: cross Abstract: Egocentric RGB-D videos offer a natural source of human dexterous manipulation demonstrations, but existing data is difficult to use for robot learnin

EgoTactile: Learning Grasp Pressure for Everyday Objects from Egocentric Video

Model ReleasesDGX agent

arXiv:2606.09243v1 Announce Type: cross Abstract: Estimating full-hand grasp pressure from egocentric video is critical for immersive VR and robotic manipulation, yet dense tactile sensing often relie

EinSort: Sorting is All We Need for Tensorizing LLM

ResearchDGX agent

arXiv:2606.08565v1 Announce Type: cross Abstract: Tensor networks provide efficient representations for compressing large neural networks. By carefully designing shapes and topologies, they can signif


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Emergence of Context Characteristics Sensitivity in Large Language Models

TutorialsDGX agent

arXiv:2606.09525v1 Announce Type: cross Abstract: During instruction fine-tuning (IFT), large language models (LLMs) learn to follow instructions by using the provided context to answer a query. While

Emergence via Phase Transitions: Mechanism Landscapes and Universal Convergence Across Complex Systems

ResearchDGX agent

arXiv:2606.07563v1 Announce Type: cross Abstract: Across machine learning, biology, and physics, independently evolving systems often converge toward strikingly similar high-level structures despite r

Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy

Model ReleasesDGX agent

arXiv:2606.08367v1 Announce Type: cross Abstract: Most evaluations of LLM agents look like exams: a discrete task, a clean environment, a score in minutes or hours. We argue that this approach is mism

Emergent alignment and the projectability of ethical personas

SafetyDGX agent

arXiv:2606.09475v1 Announce Type: new Abstract: Work on `emergent misalignment' shows that finetuning LLMs on narrow tasks can induce broadly misaligned behavior. This supports the `persona selection'

Enabling KV Caching of Shared Prefix for Diffusion Language Models

ResearchDGX agent

arXiv:2606.07571v1 Announce Type: cross Abstract: Key-value (KV) caching for shared prefixes is essential for high-throughput large language model (LLM) serving, but it faces critical challenges in em

End-to-End Context Compression at Scale

Model ReleasesDGX agent

arXiv:2606.09659v1 Announce Type: cross Abstract: Long-context language model inference is bottlenecked by memory, as the KV cache grows with context length. Recent techniques to compress the KV cache

End-to-End Training for Discrete Token LLM based TTS System

Model ReleasesDGX agent

arXiv:2606.09234v1 Announce Type: cross Abstract: Recent state-of-the-art (SOTA) text-to-speech (TTS) systems typically adopt a cascaded pipeline consisting of a speech tokenizer, an autoregressive la

Engagement Process: Rethinking the Temporal Interface of Action and Observation

AgentsDGX agent

arXiv:2605.11484v2 Announce Type: replace Abstract: Task completion in digital and physical environments increasingly involves complex temporal interaction, where actions and observations unfold over

Enhancing AI Interpretability and Safety through Localised Architectures

SafetyDGX agent

arXiv:2606.07998v1 Announce Type: cross Abstract: Recent advances in generative AI, especially powerful Large Language Models (LLMs) and Large Reasoning Models (LRMs), raise concerns over the interpre

ePC: Fast and Deep Predictive Coding in Digital Simulation

ResearchDGX agent

arXiv:2505.20137v5 Announce Type: replace-cross Abstract: Predictive Coding (PC) offers a brain-inspired alternative to backpropagation for neural network training, described as a physical system mini

EssentialGIN: a new approach for gene essentiality prediction based on graph isomorphism neural networks

ResearchDGX agent

arXiv:2606.07700v1 Announce Type: cross Abstract: Background: Prediction of essential genes (proteins), is a basic and challenging problem but at the same time very costly and time-consuming in wet-la

Evaluating Advanced Prompting on Gemini Flash for Multi-Hop Biomedical QA

Model ReleasesDGX agent

arXiv:2606.07548v1 Announce Type: cross Abstract: The MedHopQA challenge presents a critical test for Large Language Models (LLMs): complex, multi-hop reasoning in the high-stakes biomedical domain. T

Evaluating AI Investment Strategies

SafetyDGX agent

arXiv:2606.08791v1 Announce Type: cross Abstract: We study the problem of auditing a black-box algorithmic decision-maker from observable inputs and outputs alone. Our main result is an exact decompos

Evaluating Hallucinations in Domain-Adapted Large Language Models

Model ReleasesDGX agent

arXiv:2606.07521v1 Announce Type: cross Abstract: This study investigates the phenomenon of hallucinations in domain-adapted Large Language Models (LLMs), focusing on the fine-tuning of the Llama-2 mo

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting

Model ReleasesDGX agent

arXiv:2606.09809v1 Announce Type: new Abstract: AI evaluation results are produced at scale but reported inconsistently across leaderboards, model cards, benchmark papers, and company blogs. The cost

EvoCSFL: Surrogate-Assisted Evolutionary Client Selection for Efficient and Robust Federated Learning

ResearchDGX agent

arXiv:2606.07702v1 Announce Type: cross Abstract: The heterogeneity of client data and systems makes it difficult to achieve satisfactory convergence speed and robustness in federated learning with ra

Executable World Models for ARC-AGI-3 in the Era of Coding Agents

Model ReleasesDGX agent

arXiv:2605.05138v2 Announce Type: replace Abstract: We evaluate an initial coding-agent system for ARC-AGI-3 in which the agent maintains an executable Python world model, verifies it against previous

Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory

Model ReleasesDGX agent

arXiv:2606.09365v1 Announce Type: new Abstract: Medical agent systems are increasingly expected to support interactive clinical decision making rather than only static question answering. In such sett

Explaining Black-Box Language Models: Learning to Optimize Linguistically-Structured Word Subsets

Model ReleasesDGX agent

arXiv:2606.08497v1 Announce Type: new Abstract: As deep language models (DLMs) are increasingly deployed in high-stakes domains such as healthcare, understanding their decision rationale becomes param

Explaining Data Mixing Scaling Laws

TutorialsDGX agent

arXiv:2606.08167v1 Announce Type: cross Abstract: Recent research has established empirical scaling laws to predict model performance on multi-domain data mixtures. However, a theoretical understandin

Exploring the Effect of Basis Rotation on NQS Performance

Model ReleasesDGX agent

arXiv:2512.17893v2 Announce Type: replace-cross Abstract: Neural Quantum States (NQS) are powerful variational representations of quantum many-body wavefunctions, yet their performance depends sensiti

Extending Ontologies: From Dense Embeddings to Hybrid Quantum-Fuzzy Systems

ResearchDGX agent

arXiv:2606.08658v1 Announce Type: new Abstract: LLMs have revolutionized knowledge representation and retrieval, but lack the explicit modeling that knowledge ontologies possess. This paper surveys th

Eyes All Around: Design and Analysis of 360-Degree LiDAR Perception Using Equivariant Feature Learning in Unstructured Traffic

AgentsDGX agent

arXiv:2606.07626v1 Announce Type: cross Abstract: Perception in dense, unstructured urban traffic remains a major challenge for autonomous driving because of the wide variety of road users, frequent o

FADTI: Fourier and Attention Driven Diffusion for Multivariate Time Series Imputation

SafetyDGX agent

arXiv:2512.15116v2 Announce Type: replace-cross Abstract: Multivariate time series imputation is fundamental in applications such as healthcare, traffic forecasting, and biological modeling, where sen

Failure-Aware Refinement of Vision-Language Model for Lithography Defect Detection

ResearchDGX agent

arXiv:2606.08908v1 Announce Type: cross Abstract: Semiconductor lithography inspection requires reliable detection of small pattern defects such as bridge, burr, pinch, and contamination. In this stud

Failure by Interference: Language Models Make Balanced Parentheses Errors When Faulty Mechanisms Overshadow Sound Ones

ResearchDGX agent

arXiv:2507.00322v2 Announce Type: replace-cross Abstract: Despite remarkable advances in coding capabilities, language models (LMs) still struggle with simple syntactic tasks such as generating balanc

FAME: Forecastability-Aware Mixture of Experts for Heterogeneous Time Series Forecasting

SafetyDGX agent

arXiv:2606.08896v1 Announce Type: new Abstract: Large-scale retail and industrial forecasting systems contain many heterogeneous time series whose lifecycle, sparsity, volatility, seasonality, spectra

FASE: Fast Adaptive Semantic Entropy for Code Quality

AgentsDGX agent

arXiv:2606.09800v1 Announce Type: cross Abstract: Multi-agent code generation offers a promising paradigm for autonomous software development by simulating the human software engineering lifecycle. Ho

Fast LLM-Based Semantic Filtering: From a Unified Framework to an Adaptive Two-Phase Method

SafetyDGX agent

arXiv:2606.08090v1 Announce Type: cross Abstract: Evaluating a natural-language yes/no predicate over a document corpus under an accuracy target - the semantic filter - is a cornerstone of LLM-based d

Few-shot Class-variable Incremental Audio Classification via Prototype Adaptation and Pseudo Class-variable Training

ResearchDGX agent

arXiv:2606.08898v1 Announce Type: cross Abstract: In the task of few-shot class-incremental audio classification, the number of classes is assumed to always increase without considering the possibilit

FF-JEPA: Long-Horizon Planning in World Models with Latent Planners

ApplicationsDGX agent

arXiv:2606.09311v1 Announce Type: new Abstract: Joint Embedding Predictive Architectures (JEPAs) have shown promising world modeling capabilities, enabling planning in latent space by optimizing actio

FiberTune: Preserving Action-Fiber Visual Residuals in Vision-Language-Action Fine-Tuning

Local AiDGX agent

arXiv:2606.08653v1 Announce Type: cross Abstract: Action-supervised fine-tuning of vision-language-action (VLA) policies fits demonstrations effectively but constrains only the directions that change

FineGen: A VLM-based Multi-Agent Framework for Fine-Grained Image-Text Dataset Construction

Model ReleasesDGX agent

arXiv:2606.07645v1 Announce Type: cross Abstract: The scarcity of hard negative samples in current vision-language datasets significantly hinders fine-grained perception. To address this, we propose F

FIT-Print: Towards False-claim-resistant Model Ownership Verification via Targeted Fingerprint

Model ReleasesDGX agent

arXiv:2501.15509v5 Announce Type: replace-cross Abstract: Model fingerprinting has emerged as a crucial mechanism for safeguarding the intellectual property of open-source models, offering a non-intru

FlashCP: Load-Balanced Communication-Efficient Context Parallelism for LLM Training

ResearchDGX agent

arXiv:2606.08476v1 Announce Type: cross Abstract: Context parallelism (CP) is essential for training large-scale, long-context language models, as it partitions sequences to reduce memory overhead. Ho

FlashMemory-DeepSeek-V4: Lightning Index Ultra-Long Context via Lookahead Sparse Attention

Model ReleasesDGX agent

arXiv:2606.09079v1 Announce Type: cross Abstract: Conventional LLMs keep the full KV cache loaded during decoding, causing a severe GPU memory bottleneck for ultra-long context serving. In this report

FMplex: Model Virtualization for Serving Extensible Foundation Models

ApplicationsDGX agent

arXiv:2606.09643v1 Announce Type: cross Abstract: Foundation models (FMs) are increasingly used as backbones for downstream tasks across language, vision, time-series, and multimodal applications. Yet

Frequency-based Constrained Sampling for Interval Patterns

ResearchDGX agent

arXiv:2606.09666v1 Announce Type: new Abstract: Output space pattern sampling is a powerful alternative to exhaustive pattern mining for exploring large pattern spaces, as it enables users to focus on

Frequency-Domain Latent Attention Gating for Cross-Domain Token Aggregation

Model ReleasesDGX agent

arXiv:2606.08191v1 Announce Type: cross Abstract: Token aggregation is a common bottleneck in models that map token representations to sample-level predictions, yet most pooling methods operate only i

From 0-to-1 to 1-to-N: Reproducible Engineering Evidence for MetaAI Recursive Self-Design

AgentsDGX agent

arXiv:2606.09663v1 Announce Type: new Abstract: Recursive self-design refers to AI-assisted modification of the mechanisms by which an AI system is built, evaluated, and improved. This paper treats Me

From Coarse to Fine: Managing Temporal Granularity in Spatio-Temporal Data for Fine-Grained Traffic Prediction

Model ReleasesDGX agent

arXiv:2606.09392v1 Announce Type: new Abstract: Efficient acquisition, storage, and utilization of traffic data are critical challenges in spatio-temporal data management. Most traffic data systems co

From Conflict to Consensus: Boosting Medical Reasoning via Multi-Round Agentic RAG

AgentsDGX agent

arXiv:2603.03292v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) exhibit high reasoning capacity in medical question-answering, but their tendency to produce hallucinations and o

From Human Guidance to Autonomy: Agent Skill System for End-to-End LLM Deployment on Spatial NPUs

Model ReleasesDGX agent

arXiv:2606.07586v1 Announce Type: cross Abstract: Spatial neural processing units (NPUs) provide an energy-efficient platform for edge LLM inference, but efficiently deploying an LLM end-to-end on suc

From `May' to `Is': Certainty Distortion in Language Model Rewriting

Model ReleasesDGX agent

arXiv:2606.07951v1 Announce Type: cross Abstract: Humans increasingly turn to Language Models (LMs) in ways that shape beliefs and drive decisions, including discussing, rewriting, and summarizing inf

From Rigid to Dynamic: Entropy-Guided Adaptive Inference for Long-Context LLMs

Model ReleasesDGX agent

arXiv:2606.09508v1 Announce Type: new Abstract: Existing sparse attention and KV cache compression methods for long-context LLM inference typically apply fixed sparsity patterns or uniform budgets acr

From Statute to Control Flow: Span-Grounded Deontic Trees for Defeasible Scope Parsing

Model ReleasesDGX agent

arXiv:2606.08932v1 Announce Type: cross Abstract: Rule-following agents tasked with executing policies and regulations often fail via Silent Scope Omission (SSO): a model applies a general rule but si

From USD Scenes to Knowledge Graphs: Zero-Shot Ontology Grounding with LLMs

ResearchDGX agent

arXiv:2606.09134v1 Announce Type: cross Abstract: Constructing knowledge graphs from 3D simulation scenes is essential for robot task reasoning, but the key bottleneck, grounding scene objects to form

From Validator Selection to Portfolio Collection Optimization in Proof-of-Stake Blockchains

ResearchDGX agent

arXiv:2606.08282v1 Announce Type: new Abstract: We consider a problem arising in proof-of-stake blockchain environments, where agents called nominators select validators - entities responsible for mai

FunctionEvolve: Structure-Guided Symbolic Regression with LLMs

Model ReleasesDGX agent

arXiv:2606.07704v1 Announce Type: cross Abstract: Symbolic regression aims to uncover explicit scientific laws from data. Recent methods use LLMs to guide mutation from background text, which is more

FuseFSS: Efficient Secure LLM Inference with Function Secret Sharing

HardwareDGX agent

arXiv:2606.09551v1 Announce Type: cross Abstract: Two-server secure inference allows a client to query a hosted large language model (LLM) without revealing prompts or embeddings. Recent GPU systems b

GEAR-VLA: Learning Geometry-Aware Action Representations for Generalizable Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.08530v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models achieve strong benchmark performance but still struggle in real-world deployment with unseen objects, background s

Generation Properties of Stochastic Interpolation under Finite Training Set

ResearchDGX agent

arXiv:2509.21925v2 Announce Type: replace-cross Abstract: This paper investigates the theoretical behavior of generative models under finite training populations. Within the stochastic interpolation g

Generative Frontier Planning for Adaptive Peer-Referral Recruitment under Covariate-Dependent Arrivals

ResearchDGX agent

arXiv:2606.08360v1 Announce Type: cross Abstract: Peer-referral recruitment systems such as respondent-driven sampling are critical for studying and intervening on hidden populations affected by infec

Generative Reasoning Re-ranker

SafetyDGX agent

arXiv:2602.07774v5 Announce Type: replace-cross Abstract: Recent studies increasingly explore Large Language Models (LLMs) as a new paradigm for recommendation systems due to their scalability and wor

GenTSE: Enhancing Target Speaker Extraction via a Coarse-to-Fine Generative Language Model

SafetyDGX agent

arXiv:2512.20978v2 Announce Type: replace-cross Abstract: Language Model (LM)-based generative modeling has emerged as a promising direction for TSE, offering potential for improved generalization and

GIFT: LLM-Guided State-Reward Interface for Financial Reinforcement Learning

AgentsDGX agent

arXiv:2606.08450v1 Announce Type: new Abstract: Financial portfolio trading is naturally formulated as a reinforcement learning problem, where an agent sequentially rebalances assets under changing ma

GIScholarBench: Benchmarking LLM Overconfidence in GIS Research

Model ReleasesDGX agent

arXiv:2606.08036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in academic research workflows, but scholarly tasks require high factual precision and therefore ex

← Previous
1…143144145146147…358
Next →