AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
9 Jun 2026

GlobeAudio: A Multilingual Multicultural Benchmark for Naturalistic Evaluation of Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2606.08194v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) integrate audio perception and language understanding within a unified framework, enabling a wide range of real-wo

Goal-Oriented Reasoning for RAG-based Memory in Conversational Agentic LLM Systems

AgentsDGX agent

arXiv:2605.12213v2 Announce Type: replace Abstract: LLM-based conversational AI agents struggle to maintain coherent behavior over long horizons due to limited context. While RAG-based approaches are

Governance Controls for AI-Generated Test Artifacts in Autonomous Software Testing

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.08806v1 Announce Type: cross Abstract: Artificial Intelligence (AI) and Large Language Models (LLMs) are increasingly used in autonomous software testing; however, AI-generated test artifac

Graph-to-SFILES: Control structure prediction from process topologies using generative artificial intelligence

ResearchDGX agent

arXiv:2412.00508v2 Announce Type: replace-cross Abstract: Control structure design is an important but tedious step in P&ID development. Generative artificial intelligence (AI) promises to reduce P&ID

Graph2Idea:Retrieval-Augmented Scientific Idea Generation with Graph-Structured Contexts

Model ReleasesDGX agent

arXiv:2606.09105v1 Announce Type: new Abstract: Generating novel, feasible, and high-quality research ideas is an important yet challenging task in scientific discovery.Recent Large Language Model (LL

GraphLoRA: Structure-Aware Low-Rank Adaptation for Large Language Model Recommendation

Model ReleasesDGX agent

arXiv:2606.07526v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown strong potential for recommendation (LLMRec) due to their powerful reasoning and generalization abilities. How

GVC-Seg: Training-Free 3D Instance Segmentation via Geometric Visual Correspondence

SafetyDGX agent

arXiv:2606.08014v1 Announce Type: cross Abstract: Accurate 3D instance segmentation in point cloud data is critical for machine vision applications. Recent advancements leverage multiple pre-trained f

HA-VLN 2.0: An Open Benchmark and Leaderboard for Human-Aware Navigation in Discrete and Continuous Environments with Dynamic Multi-Human Interactions

Model ReleasesDGX agent

arXiv:2503.14229v4 Announce Type: replace Abstract: Vision-and-Language Navigation (VLN) has been studied mainly in either discrete or continuous spaces, with little attention to dynamic, crowded envi

Hacking Generative Perplexity: Why Unconditional Text Evaluation Needs Distributional Metrics

Model ReleasesDGX agent

arXiv:2606.08417v1 Announce Type: cross Abstract: Diffusion and continuous flow-based language models have emerged as the leading non-autoregressive alternatives to language modeling. Progress in both

HARBOR: A Harness Framework for Agentic Robot Reinforcement Learning

SafetyDGX agent

arXiv:2606.08610v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a powerful paradigm for robot learning, particularly in sim-to-real settings, but its broader adoption remains

Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops

Model ReleasesDGX agent

arXiv:2606.08960v1 Announce Type: cross Abstract: Agent benchmarks score submissions with outcome verifiers that are typically hand-written and brittle, leaving them open to reward hacking. We audit 1

Harmonia: End-to-End RAG Serving Optimization

ResearchDGX agent

arXiv:2505.07833v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) improves the reliability of large language models by integrating external knowledge, but serving RAG pipe

Harness Engineering for Physical AI: Robot Middleware Is the Harness Layer

SafetyDGX agent

arXiv:2606.09416v1 Announce Type: cross Abstract: Robot middleware faces a new role in the era of Physical AI. Learned policies, planners, and vision-language-action (VLA) models now enter deployed ro

HARP: Efficient Data Selection for Finetuning Large Language Models

SafetyDGX agent

arXiv:2606.07690v1 Announce Type: cross Abstract: Finetuning data selection requires balancing two competing goals: selecting examples that improve the downstream objective, and doing so without repea

HASA: Subnet Allocation for Compute-Constrained Model-Heterogeneous Federated Learning

Model ReleasesDGX agent

arXiv:2606.07621v1 Announce Type: cross Abstract: Edge services increasingly use federated learning to personalize on-device models while keeping sensitive data local. In practice, deployments must ha

Hiding in Plain Floats: Steganographic Carriers for Indirect Prompt and Content Injection

ResearchDGX agent

arXiv:2606.08403v1 Announce Type: cross Abstract: Text-centered prompt-injection defenses assume that the malicious signal is visible in one of the inspected text views. We study a reproducible LLM01-

How Context Shapes Truth: Geometric Transformations of Statement-level Truth Representations in LLMs

ResearchDGX agent

arXiv:2601.06599v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) often encode whether a statement is true as a vector in their residual stream activations. These vectors, also kn

How Deep Are Deep GPs, Really? A Sharp Threshold and a Non-Gaussian Limit for Compositional GPs

ResearchDGX agent

arXiv:2606.08218v1 Announce Type: cross Abstract: Compositional priors describe the generic properties of layered functions in deep Bayesian models, where deep neural networks with random weights are

How Hyper-Datafication Impacts the Sustainability Costs in Frontier AI

ResearchDGX agent

arXiv:2602.00056v4 Announce Type: replace-cross Abstract: Large-scale data has fuelled the success of frontier artificial intelligence (AI) models over the past decade. This expansion has relied on su

How Many Counterfactuals Does It Take? Probing VLM Hallucinations Through Circuits and Causal Effects

ResearchDGX agent

arXiv:2606.08777v1 Announce Type: cross Abstract: Visual Language Models (VLMs) are known to produce hallucinated predictions that are not grounded in visual evidence, yet existing approaches lack a p

How Much Dense Attention is Necessary? Oracle-Guided Sparse Prefill for Full/GQA Layers in Hybrid Long-Context Models

Model ReleasesDGX agent

arXiv:2606.07703v1 Announce Type: cross Abstract: Long-context prefill remains expensive because full/GQA layers still score the historical sequence, even in hybrid models with local, sparse, linear,

How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions

Model ReleasesDGX agent

arXiv:2606.08051v1 Announce Type: new Abstract: Financial transaction processing requires extracting structured merchant information from noisy, abbreviated bank transaction strings at scale. Our curr

How Transformers Reject Wrong Answers: Rotational Dynamics of Factual Constraint Processing

ResearchDGX agent

arXiv:2603.13259v2 Announce Type: replace-cross Abstract: When a decoder-only transformer is forced to process matched correct and incorrect single-token continuations of a factual query, the two path

Human-Centered Benchmarking of Driver Monitoring Models

SafetyDGX agent

arXiv:2606.08123v1 Announce Type: cross Abstract: Vision-based driver monitoring systems are increasingly deployed in safety-critical intelligent transportation settings, yet they are almost always co

Hybrid E-Assessment in Higher Education: Semi-Automated Grading of Paper-Based Written Examinations

SafetyDGX agent

arXiv:2606.08855v1 Announce Type: new Abstract: This paper examines the limitations of fully digital and partially digital e-assessment approaches in summative examinations in higher education. The an

Hybrid Neural Network and Conventional Controller Approach for Robust Control of Highly Unstable Systems: Application to Tilt-Rotor Control

ApplicationsDGX agent

arXiv:2606.08714v1 Announce Type: cross Abstract: Multirotors are widely used in applications ranging from surveillance to precision agriculture, yet conventional designs remain limited by their under

Hybrid Robustness Verification for Spatio-Temporal Neural Networks

Model ReleasesDGX agent

arXiv:2606.09746v1 Announce Type: cross Abstract: With AI increasingly deployed in safety-critical systems, providing formal robustness guarantees for the underlying models is essential. Existing veri

Hybridizing Equilibrium Propagation with Ising Machines for Efficient Energy-Based Learning

HardwareDGX agent

arXiv:2606.09112v1 Announce Type: cross Abstract: The rapid evolution of artificial intelligence has led to substantial advances in deep neural networks. Nonetheless, conventional GPU-based training r

Hyperflux: Pruning Reveals Importance

ResearchDGX agent

arXiv:2504.05349v4 Announce Type: replace-cross Abstract: Network pruning is used to reduce inference latency and power consumption in large neural networks. However, most methods focus on empirical r

I-Segmenter: Integer-Only Vision Transformer for Efficient Semantic Segmentation

ApplicationsDGX agent

arXiv:2509.10334v2 Announce Type: replace-cross Abstract: Vision Transformers (ViTs) have recently achieved strong results in semantic segmentation, yet their deployment on resource-constrained device

'I understand your perspective': LLM Persuasion and Sycophancy through the Lens of Communicative Action Theory

ResearchDGX agent

arXiv:2606.08076v1 Announce Type: cross Abstract: Large Language Models (LLMs) can generate high-quality arguments, yet their ability to engage in nuanced and persuasive communicative actions remains

I Was Scrolling and Then I Saw a Pregnant Strawberry

ResearchDGX agent

arXiv:2606.09589v1 Announce Type: cross Abstract: AI minidramas (also known as fruit dramas) are short, algorithmically distributed generative AI video series featuring anthropomorphized characters th

IDEQ -- Improving Diffusion Models for the Traveling Salesman Problem (TSP) by Leveraging the Structure of the Solution Space

Model ReleasesDGX agent

arXiv:2412.13858v2 Announce Type: replace Abstract: We investigate diffusion models to solve the Traveling Salesman Problem. Building on the recent DIFUSCO and T2TCO approaches, we propose IDEQ. IDEQ

IEA: Amateur-Friendly Conversational Image Editing Agent via Three Stages of Multitask Alignment

SafetyDGX agent

arXiv:2606.08016v1 Announce Type: cross Abstract: Current image editing software often hinges on fixed filters or expert tuning, leaving a gap between amateur users' intent and outcomes. Creations by

Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation

ResearchDGX agent

arXiv:2606.07936v1 Announce Type: cross Abstract: Human evaluation plays a critical role in assessing the quality of generated text. However, the reliability and reproducibility of these evaluations d

Impacts of Histories and Models on LLM Grading: A Study in Advanced Software Engineering Courses

SafetyDGX agent

arXiv:2606.08400v1 Announce Type: cross Abstract: Graduate-level research reading report assessment creates a substantial labor burden for educators. While large language models (LLMs) hold great pote

Implementing Grassroots Logic Programs with Multiagent Transition Systems and AI (Full Version)

Model ReleasesDGX agent

arXiv:2602.06934v4 Announce Type: replace-cross Abstract: Grassroots Logic Programs (GLP) is a concurrent logic programming language in which logic variables are partitioned into paired readers and wr

Implicit Causal Graph Construction in Text via Chain Discovery

ResearchDGX agent

arXiv:2606.07525v1 Announce Type: cross Abstract: Causal graphs in text are typically populated by observable, predefined events. In contrast, we study implicit causal graph construction from text by

Improving Multimodal Reasoning via Worst Dimension Optimization

ResearchDGX agent

arXiv:2606.07801v1 Announce Type: new Abstract: Multimodal reasoning requires a path that retains integrity over a wide range of constraints, from visual grounding to logic consistency. However, the c

IMUG-Bench: Benchmarking Unified Multimodal Models on Interleaved Understanding and Generation

Model ReleasesDGX agent

arXiv:2606.09169v1 Announce Type: new Abstract: In recent years, unified multimodal models (UMMs) have emerged to support both understanding and generation within a single framework. Mastering dynamic

In-Context Reinforcement Learning via Communicative World Models

AgentsDGX agent

arXiv:2508.06659v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) agents often struggle to generalize to new tasks and contexts without updating their parameters, mainly because th

InA-Probe: Instruction-Aware Active Probing for Time Series Forecasting with LLMs

SafetyDGX agent

arXiv:2606.08601v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently demonstrated impressive potential for time series forecasting. However, existing methods predominantly rely o

Inference-Time Conformal Reasoning with Valid Factuality Control for Large Language Models

ResearchDGX agent

arXiv:2606.08831v1 Announce Type: new Abstract: Large language models (LLMs) increasingly perform multi-step reasoning, where intermediate claims form implicit directed acyclic graphs whose node corre

INFUSER: Influence-Guided Self-Evolution Improves Reasoning

ResearchDGX agent

arXiv:2606.09052v1 Announce Type: cross Abstract: Self-evolution offers a scalable path to stronger reasoning: a pretrained language model improves itself with only minimal external supervision. Yet e

Instrumental convergence and power-seeking

SafetyDGX agent

arXiv:2606.08832v1 Announce Type: new Abstract: Recent years have seen increasing concern that artificial intelligence may soon pose an existential risk to humanity. One leading ground for concern is

Instrumented data for causal scientific machine learning

ResearchDGX agent

arXiv:2606.07865v1 Announce Type: cross Abstract: Scientific machine learning is limited less by model size than by the data it is trained on. Observational data records what happened but not why; tem

Integrating Deep Learning Demand Forecasting with Multi-Objective Optimization for Circular Coffee Supply Chains: A Data-Driven Framework for Cost, Emissions, and Freshness Management

Model ReleasesDGX agent

arXiv:2606.08314v1 Announce Type: new Abstract: The coffee supply chain is one of the most complex agri-food networks, marked by geographically dispersed production, multi-tier coordination, and high

Intelligent Character Recognition of Handwritten Forms with Deep Neural Networks

ResearchDGX agent

arXiv:2606.08858v1 Announce Type: cross Abstract: The automatic processing of handwritten forms remains a challenging task, wherein detection and subsequent classification of handwritten characters ar

Internalizing Geometric Law: Learning from Solver Residuals for Precision-Critical Generation

Model ReleasesDGX agent

arXiv:2606.09278v1 Announce Type: cross Abstract: Large Language Models frequently hallucinate in precision-critical domains such as technical diagramming and mechanical design, where outputs must sat

Intrinsic Selection and Particle Resampling for Inference-Time Scaling Beyond Domain Verifiability

ResearchDGX agent

arXiv:2606.08850v1 Announce Type: cross Abstract: Inference-Time Scaling (ITS) has largely succeeded in verifiable domains like math and coding, where cheap verification enables scalable output select

Investigating the Histogram Loss in Regression

ResearchDGX agent

arXiv:2402.13425v3 Announce Type: replace-cross Abstract: It is becoming increasingly common in regression to train neural networks that model the entire distribution even if only the mean is required

IRAM-Omega-Q: A Computational Framework for Uncertainty Regulation in Adaptive Agents

ResearchDGX agent

arXiv:2603.16020v2 Announce Type: replace Abstract: Adaptive agents operating under uncertainty must do more than optimize task outputs: they must maintain a workable internal state under noise, pertu

Item Response Scaling Laws: A Measurement Theory Approach for Efficient and Generalizable Neural Scaling Estimation

Model ReleasesDGX agent

arXiv:2606.07616v1 Announce Type: cross Abstract: Scaling laws provide a fundamental framework for understanding the performance of Language Models (LMs), yet deriving them requires prohibitively expe

Jas: AI-Paired Engineering as a Revival of N-Version Programming

ApplicationsDGX agent

arXiv:2606.07828v1 Announce Type: cross Abstract: I report a case study in AI-paired software engineering: five working ports of a vector illustration application across Rust, Swift, OCaml, Python, an

Joint Structural Pruning and Mixed-Precision Quantization for LLM Compression

ResearchDGX agent

arXiv:2606.07819v1 Announce Type: new Abstract: Recently, the efficiency of Large Language Models (LLMs) deployment has become a critical concern in practical applications. While post-training quantiz

Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models

SafetyDGX agent

arXiv:2602.12996v2 Announce Type: replace-cross Abstract: Knowledge augmentation has significantly enhanced the performance of Large Language Models (LLMs) in knowledge-intensive tasks. However, exist

Knowledge Graphs and Reasoning LLMs for Finding Simple Yet Effective Transcriptomic Perturbation Predictors

ResearchDGX agent

arXiv:2606.08816v1 Announce Type: cross Abstract: Predicting the effect of an unseen gene knockout perturbation on transcriptomic gene expression remains a highly challenging problem for virtual cell

Knowledge-Inclusive Adaptive Physics-Informed Neural Network for Microbial Interaction Modelling

Model ReleasesDGX agent

arXiv:2606.07686v1 Announce Type: cross Abstract: Physics-Informed Neural Network (PINN) is a way of including knowledge in the form of equations in Machine Learning methods. Beyond equations, knowled

Kunlun: Establishing Scaling Laws for Massive-Scale Recommendation Systems through Unified Architecture Design

HardwareDGX agent

arXiv:2602.10016v3 Announce Type: replace-cross Abstract: Deriving predictable scaling laws that govern the relationship between model performance and computational investment is crucial for designing

Land cover and flood type govern the detection limits of satellite-based flood mapping across diverse global flood events

ApplicationsDGX agent

arXiv:2606.07780v1 Announce Type: new Abstract: Floods are among the most destructive natural hazards, and their increasing frequency under climate change makes satellite-based inundation mapping esse

← Previous
1…144145146147148…358
Next →