AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
4 Jun 2026

Learning Empirically Admissible Neural Heuristics for Combinatorial Search

SafetyDGX agent

arXiv:2606.04860v1 Announce Type: cross Abstract: Finding optimal solution paths for combinatorial puzzles like the Rubik's Cube, sliding tile puzzles, and Lights Out remains a classical challenge in

Learning Long Range Spatio-Temporal Representations over Continuous Time Dynamic Graphs with State Space Models

Model ReleasesDGX agent

arXiv:2606.04672v1 Announce Type: cross Abstract: Continuous-time dynamic graphs (CTDGs) provide a richer framework to capture fine-grained temporal patterns in evolving relational data. Long-range in

Learning While Acting: A Skill-Enhanced Test-Time Co-Evolution Framework for Online Lifelong Learning Agents

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.04815v1 Announce Type: cross Abstract: Lifelong learning is essential for Large Language Model (LLM) agents operating in dynamic, interactive environments. However, existing lifelong learni

LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection

Model ReleasesDGX agent

arXiv:2606.04050v1 Announce Type: cross Abstract: Existing quantization methods are fundamentally limited by rigid, integer-based bit-widths (e.g., 2, 3-bit), resulting in a ``deployment gap' where La

LLM Compression with Jointly Optimizing Architectural and Quantization choices

HardwareDGX agent

arXiv:2606.04063v1 Announce Type: cross Abstract: Deploying large language models (LLMs) is challenging due to their significant memory and computational requirements. While some methods address this

Longer Context, Deeper Thinking: Uncovering the Role of Long-Context Ability in Reasoning

Model ReleasesDGX agent

arXiv:2505.17315v2 Announce Type: replace Abstract: Recent language models exhibit strong reasoning capabilities, yet the influence of long-context capacity on reasoning remains underexplored. In this

LoopMoE: Unifying Iterative Computation with Mixture-of-Experts for Language Modeling

Model ReleasesDGX agent

arXiv:2606.04438v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) and looped architectures scale models along two orthogonal axes, namely parameter capacity and effective depth. However, main

Low-Rank Decay for Grokking in Scale-Invariant Transformers: A Spectral-Geometric View

ResearchDGX agent

arXiv:2606.04405v1 Announce Type: cross Abstract: Modern Transformer architectures frequently employ normalization mechanisms such as RMSNorm and Query-Key Normalization, making parts of the model app

M^3Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks

Model ReleasesDGX agent

arXiv:2606.05008v1 Announce Type: cross Abstract: As multi-modal models advance towards long-form video understanding, memory emerges as a critical capability. Despite substantial efforts in developin

Making Expert Reasoning Learnable with Self-Distillation

ResearchDGX agent

arXiv:2602.02405v2 Announce Type: replace-cross Abstract: Improving the reasoning capabilities of large language models (LLMs) typically relies either on the model's ability to sample a correct soluti

MapAgent: An Industrial-Grade Agentic Framework for City-scale Lane-level Map Generation

AgentsDGX agent

arXiv:2606.04513v1 Announce Type: new Abstract: Lane-level maps are critical infrastructure for autonomous driving and lane-level navigation, yet constructing and maintaining standardized lane network

MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models

SafetyDGX agent

arXiv:2606.04027v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising partially masked sequences under bidirectional context, exposing a safe

Measuring What Matters: Synthetic Benchmarks for Concept Bottleneck Models

TutorialsDGX agent

arXiv:2606.04326v1 Announce Type: cross Abstract: Concept bottleneck models predict outcomes from high-level concepts detected in inputs. Although concepts provide a simple way to reap benefits from i

MedForge: Interpretable Medical Deepfake Detection via Forgery-aware Reasoning

Model ReleasesDGX agent

arXiv:2603.18577v2 Announce Type: replace Abstract: Text-guided image editors can now manipulate authentic medical scans with high fidelity, enabling lesion implantation/removal that threatens clinica

MemoryDocDataSet: A Benchmark for Joint Conversational Memory and Long Document Reasoning

Model ReleasesDGX agent

arXiv:2606.04442v1 Announce Type: cross Abstract: AI systems increasingly need to combine two demanding capabilities: navigating multi-session conversation history and performing deep reading comprehe

MENTOR: A Metacognition-Driven Self-Evolution Framework for Uncovering and Mitigating Implicit Domain Risks in LLMs

SafetyDGX agent

arXiv:2511.07107v3 Announce Type: replace Abstract: Ensuring the safety of Large Language Models (LLMs) is critical for real-world deployment. However, current safety measures often fail to address im

MesaNet: Sequence Modeling by Locally Optimal Test-Time Training

Model ReleasesDGX agent

arXiv:2506.05233v2 Announce Type: replace-cross Abstract: Sequence modeling is currently dominated by causal transformer architectures that use softmax self-attention. Although widely adopted, transfo

Metric-Aware Hybrid Forecasting for the CTF4Science Lorenz Challenge

Model ReleasesDGX agent

arXiv:2606.04191v1 Announce Type: cross Abstract: We describe our approach to the CTF4Science Lorenz challenge, a benchmark that mixes short-horizon forecasting, long-time distribution matching, and t

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers

ResearchDGX agent

arXiv:2601.07036v2 Announce Type: replace-cross Abstract: Hybrid reasoning language models are commonly controlled through high-level Think/No-think instructions to regulate reasoning behavior, yet we

MimeLens: Position-Agnostic Content-Type Detection for Binary Fragments

Model ReleasesDGX agent

arXiv:2606.04171v1 Announce Type: cross Abstract: File-type classification underlies many workflows like malware triage, forensic carving, packet inspection, and storage indexing. Learned systems such

MIRAGE: Mobile Agents with Implicit Reasoning and Generative World Models

AgentsDGX agent

arXiv:2606.04627v1 Announce Type: new Abstract: Mobile agents are increasingly expected to operate everyday applications from screenshots and language goals, where reliable control requires reasoning

MM-BizRAG: Rethinking Multimodal Retrieval-Augmented Generation for General Purpose Enterprise Q&A

SafetyDGX agent

arXiv:2606.04231v1 Announce Type: cross Abstract: Recent advances in multimodal retrieval-augmented generation (MM-RAG) have shifted toward minimal parsing, relying on page-level images for producing

Model-Preserving Adaptive Rounding

ResearchDGX agent

arXiv:2505.22988v3 Announce Type: replace-cross Abstract: The goal of quantization is to produce a compressed model whose output distribution is as close to the original model's as possible. To do thi

MorphoQuant: Modality-Aware Quantization for Omni-modal Large Language Models

SafetyDGX agent

arXiv:2606.04349v1 Announce Type: cross Abstract: Conventional Post-Training Quantization (PTQ) methods struggle with 4-bit Omni-modal Large Language Models (OLLMs) due to the extreme distribution het

MuCO: Generative Peptide Cyclization Empowered by Multi-stage Conformation Optimization

ResearchDGX agent

arXiv:2602.11189v2 Announce Type: replace-cross Abstract: Modeling peptide cyclization is critical for the virtual screening of candidate peptides with desirable physical and pharmaceutical properties

Multi-Column RBF Neural Network Using Adaptive and Non-Adaptive Particle Swarm Optimization

Model ReleasesDGX agent

arXiv:2606.05150v1 Announce Type: cross Abstract: The radial basis function neural network (RBFN) trained with a gradient descending algorithm provides an effective fully connected structure in both s

Multi-Granularity 3D Kidney Lesion Characterization from CT Volumes

ResearchDGX agent

arXiv:2606.04365v1 Announce Type: cross Abstract: Radiology reports describe kidney lesions by type, size, enhancement, and attenuation, yet existing 3D methods predict only at the patient or organ le

Multi-SPIN: Multi-Access Speculative Inference for Cooperative Token Generation at the Edge

Model ReleasesDGX agent

arXiv:2606.04581v1 Announce Type: cross Abstract: Speculative inference (SPIN) was originally developed as an efficient architecture to accelerate Large Language Models (LLMs). In this work, we propos

Need to Know: Contextual-Integrity-Grounded Query Rewriting for Privacy-Conscious LLM Delegation

Model ReleasesDGX agent

arXiv:2606.04067v1 Announce Type: cross Abstract: As LLMs become increasingly woven into everyday workflows, user queries sent to cloud hosted LLMs routinely mix task-essential content with task non-e

Neetyabhas: A Framework for Uncertainty-Aware Public Policy Optimization in Rational Agent-Based Models

SafetyDGX agent

arXiv:2606.04562v1 Announce Type: new Abstract: Purpose The WHO's COVID-19 non-pharmaceutical interventions (e.g., lockdowns, vaccinations) effectively curb transmission but impose heavy economic stra

Neural Radiated-Noise Fields for Unmanned Underwater Vehicle Noise Spectrum Prediction in Three-Dimensional Scenes

ResearchDGX agent

arXiv:2606.04008v1 Announce Type: cross Abstract: Radiated noise in unmanned underwater vehicles (UUVs) is an important indicator for characterizing acoustic signatures and evaluating platform perform

NoRA: Evaluating Grounded Reasonableness in Visual First-person Normative Action Reasoning

Model ReleasesDGX agent

arXiv:2606.04806v1 Announce Type: cross Abstract: LLMs and agentic systems are increasingly deployed in social environments, making normative competence critical for safe and appropriate behavior. How

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation

Model ReleasesDGX agent

arXiv:2606.04402v1 Announce Type: new Abstract: Modern reasoning models can allocate different amounts of test-time computation, such as thinking tokens, model calls, or compute budget, to different t

Notarized Agents: Receiver-Attested Confidential Receipts for AI Agent Actions

AgentsDGX agent

arXiv:2606.04193v1 Announce Type: cross Abstract: Current AI agent observability is structurally compromised: the entity producing the activity log is the same entity whose activity is being logged. A

OA-CutMix: Correcting the Label Bias of CutMix

SafetyDGX agent

arXiv:2606.04820v1 Announce Type: cross Abstract: CutMix has become the de facto standard mixing augmentation, yet its label assignment rests on a flawed assumption: The area of the pasted patch faith

OckBench: Measuring the Efficiency of LLM Reasoning

Model ReleasesDGX agent

arXiv:2511.05722v3 Announce Type: replace-cross Abstract: Large language models (LLMs) such as GPT-5 and Gemini 3 have pushed the frontier of automated reasoning and code generation. Yet current bench

On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers

SafetyDGX agent

arXiv:2603.28762v2 Announce Type: replace-cross Abstract: Modern Text-to-Image (T2I) diffusion models have achieved remarkable semantic alignment, yet they often suffer from a significant lack of vari

Online Skill Learning for Web Agents via State-Grounded Dynamic Retrieval

Model ReleasesDGX agent

arXiv:2606.04391v1 Announce Type: new Abstract: Language agents increasingly rely on reusable skills to improve multi-step web automation across related tasks. A growing line of work studies online sk

OpenRFM: Dissecting Relational In-Context Learning

ResearchDGX agent

arXiv:2606.04320v1 Announce Type: cross Abstract: Relational Foundation Models (RFMs) promise a single pre-trained predictor that, given any relational database, returns predictions in one forward pas

Optical-Guided Neural Collapse for SAR Few-Shot Class Incremental Learning

Model ReleasesDGX agent

arXiv:2606.04528v1 Announce Type: cross Abstract: Few-shot class-incremental learning (FSCIL) in synthetic aperture radar imagery presents unique challenges due to severe data scarcity and SAR-specifi

Outcome-Based RL Provably Leads Transformers to Reason, but Only With the Right Data

SafetyDGX agent

arXiv:2601.15158v4 Announce Type: replace-cross Abstract: Transformers trained via Reinforcement Learning (RL) with outcome-based supervision can spontaneously develop the ability to generate intermed

Overview of the EReL@MIR 2025 Multimodal Document Retrieval Challenge (Track 1)

ResearchDGX agent

arXiv:2606.04240v1 Announce Type: cross Abstract: Retrieval over visually-rich documents, pages that interleave text with figures, tables, and charts, is essential for multimodal retrieval-augmented g

ParetoPilot: Zero-Surrogate Offline Multi-Objective Optimization via Infer-Perturb-Guide Diffusion

TutorialsDGX agent

arXiv:2606.04468v1 Announce Type: cross Abstract: Offline multi-objective optimization (Offline MOO) aims to discover novel Pareto-optimal designs based on static datasets without expensive environmen

Parthenon Law: A Self-Evolving Legal-Agent Framework

AgentsDGX agent

arXiv:2606.04602v1 Announce Type: new Abstract: As agents grow more capable, legal-domain LLM agents promise to turn document-heavy matters into reviewable work products -- yet reliable deployment fac

PerceptTwin: Semantic Scene Reconstruction for Iterative LLM Planning and Verification

SafetyDGX agent

arXiv:2606.04226v1 Announce Type: cross Abstract: Simulation environments are useful for both robot policy learning and planning verification and validation. Traditionally, the process of creating a s

PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?

Model ReleasesDGX agent

arXiv:2602.01146v2 Announce Type: replace Abstract: Conversational assistants are increasingly integrating long-term memory with large language models (LLMs). This persistence of memories, e.g., the u

Physics-Informed Machine Learning for Short-Term Flood Prediction

SafetyDGX agent

arXiv:2606.04143v1 Announce Type: cross Abstract: Accurate flood forecasting is essential for mitigating disaster risks and protecting communities. However, purely data-driven machine learning models

Physics-Informed Neural Engine Sound Modeling with Differentiable Pulse-Train Synthesis

ResearchDGX agent

arXiv:2603.09391v2 Announce Type: replace-cross Abstract: Engine sounds originate from sequential exhaust pressure pulses rather than sustained harmonic oscillations. While neural synthesis methods ty

Plan First, Judge Later, Run Better: A DMAIC-Inspired Agentic System for Industrial Anomaly Detection

SafetyDGX agent

arXiv:2606.04599v1 Announce Type: new Abstract: Large language model (LLM) agents have shown promise in automating complex data-analysis workflows, but their reliable deployment remains challenging in

Plan, Watch, Recover: A Benchmark and Architectures for Proactive Procedural Assistance

Model ReleasesDGX agent

arXiv:2606.04970v1 Announce Type: cross Abstract: We envision a proactive multi-modal assistant system which gives users real-time step-by-step guidance on a procedural task, autonomously deciding ext

Platonic Transformers: A Solid Choice For Equivariance

ResearchDGX agent

arXiv:2510.03511v3 Announce Type: replace-cross Abstract: While widespread, Transformers lack inductive biases for geometric symmetries common in science and computer vision. Existing equivariant meth

POLARIS: Guiding Small Models to Write Long Stories

SafetyDGX agent

arXiv:2606.04095v1 Announce Type: cross Abstract: Small open-weight models struggle at long-form creative writing: their generated stories either fall far short of the requested length, or their quali

PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay

Model ReleasesDGX agent

arXiv:2603.23841v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) are increasingly used as primary sources of information, their potential for political bias may impact thei

Position: Deployed Reinforcement Learning should be Continual

AgentsDGX agent

arXiv:2606.04029v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has received increasing attention and adoption in real-world use cases. Most of these systems follow a train-then-fix para

Proof-Carrying Agent Actions: Model-Agnostic Runtime Governance for Heterogeneous Agent Systems

Model ReleasesDGX agent

arXiv:2606.04104v1 Announce Type: cross Abstract: Agent systems execute through runtimes with very different control points: local coding tools, framework SDKs, managed agent platforms, API gateways,

Provably Auditable and Safe LLM Agents from Human-Authored Ontologies

AgentsDGX agent

arXiv:2606.04903v1 Announce Type: cross Abstract: We introduce the LLM agent architecture Agentic Redux, intended for use with nontrivial problem domains that require linear auditability. Using the ty

QO-Bench: Diagnosing Query-Operator-Preserving Retrieval over Typed Event Tuples

Model ReleasesDGX agent

arXiv:2606.04646v1 Announce Type: cross Abstract: Many real-world questions over business, legal, and scientific corpora are natural-language versions of database-style queries over records latent in

Quantum entanglement provides a competitive advantage in adversarial games

Model ReleasesDGX agent

arXiv:2603.10289v2 Announce Type: replace-cross Abstract: Whether uniquely quantum resources confer advantages in fully classical, competitive environments remains an open question. Competitive zero-s

QuBLAST: A Framework for Quantizing Large Language Models with Block-Level Compression Approach and Activation Scaling Strategy

Model ReleasesDGX agent

arXiv:2606.04620v1 Announce Type: cross Abstract: LLMs have become the state-of-the-art algorithms for solving NLP tasks. However, they typically come at huge computational and memory costs, thus maki

R-APS: Compositional Reasoning and In-Context Meta-Learning for Constrained Design via Reflective Adversarial Pareto Search

Local AiDGX agent

arXiv:2606.04823v1 Announce Type: new Abstract: Large language models (LLMs) are fluent on open-ended tasks, yet in agentic settings, where a system must plan, use tools, and act over extended horizon

← Previous
1…160161162163164…358
Next →