AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

88,316Total entries
1Added by human
88,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,550 results
11 Aug 2026

Improving Generalization Robustness of Multimodal RLVR

SafetyDGX agent

arXiv:2608.08802v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) makes Multimodal Large Language Models more accurate, but the gains are brittle: simply paraphrasi

Instruction Set and Language for Hypergraphs

Model ReleasesDGX agent

arXiv:2607.10194v2 Announce Type: replace-cross Abstract: We present IsalHG, a method for representing the structure of any finite, connected hypergraph of bounded hyperedge arity as a string over a c

Introducing 𝗘𝘅𝘁𝗿𝗮𝗰𝘁𝗕𝗲𝗻𝗰𝗵: the most comprehensive benchmark for information extraction from complex enterprise documents. Our app…

Model ReleasesDGX agent

Introducing 𝗘𝘅𝘁𝗿𝗮𝗰𝘁𝗕𝗲𝗻𝗰𝗵: the most comprehensive benchmark for information extraction from complex enterprise documents. Our applied research team tested: 14 systems — frontier VLMs, coding agents, ex

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Is the ACL Responsible NLP Checklist a Box-Ticking Exercise? A Large-Scale Analysis of EMNLP 2025

Model ReleasesDGX agent

arXiv:2608.09280v1 Announce Type: new Abstract: Responsible NLP practice includes a) transparency, b) ethics, and c) societal impacts. The Responsible NLP Checklist aims to push these goals, and promo

Keep It Simple: Multi-Key Episodic Memory Retrieval for Ultra-Long Video Understanding

AgentsDGX agent

arXiv:2608.07663v1 Announce Type: cross Abstract: When videos extend from hours to days, directly processing them end-to-end becomes impractical for current Multi-modal Large Language Models (MLLMs).

Label-Free Parkinson's Disease Screening from Face and Voice through Mechanistic Interpretability

Model ReleasesDGX agent

arXiv:2608.08976v1 Announce Type: new Abstract: Parkinson's disease (PD) is the second most common neurodegenerative disorder. Typical machine learning screening methods require PD labels, but the ava

Learning from Consensus and Disagreement: Unsupervised On-Policy Self-Distillation with Minority-Trajectory Contrast

SafetyDGX agent

arXiv:2608.08764v1 Announce Type: cross Abstract: On-policy self-distillation improves language-model reasoning by querying a teacher on states actually visited by the student. Recent methods create a

Learning Structural Illumination for Unsupervised Low-light Enhancement

Model ReleasesDGX agent

arXiv:2608.08153v1 Announce Type: new Abstract: Existing unsupervised low-light image enhancement (LLIE) methods often estimate illumination directly from the entire low-light input, without separatin

Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates

SafetyDGX agent

arXiv:2608.00326v2 Announce Type: replace Abstract: Tool calling allows large language models (LLMs) to invoke external computation during problem solving, a useful capability in various fields includ

Leveraging the frontier ecosystem We work with neolabs like @trajectorylabs and @EngramLab and inference + post-training infra providers lik…

ApplicationsDGX agent

Leveraging the frontier ecosystem We work with neolabs like @trajectorylabs and @EngramLab and inference + post-training infra providers like @baseten, @FireworksAI_HQ, and @appliedcompute to post-tra

LGNNIC: Acceleration of Large-Scale GNN Training using SmartNICs

Model ReleasesDGX agent

arXiv:2608.07733v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are widely used across domains such as natural sciences, social network analysis, chip design, and recommendation systems

LIBAD: A Multimodal Anomaly Detection Benchmark for Li-Ion Battery Electrode Manufacturing

Model ReleasesDGX agent

arXiv:2608.07958v1 Announce Type: new Abstract: Multimodal industrial anomaly detection has largely focused on discrete products using strongly correlated RGB and 3D observations, leaving continuous p

LibraSpec: Dynamic Diffusion-Based Speculative Decoding via Marginal-Gain-Driven Optimization

ResearchDGX agent

arXiv:2608.08721v1 Announce Type: cross Abstract: Speculative decoding accelerates large language model inference by drafting multiple tokens for parallel verification, with efficiency critically dete

LLM-Guided Heuristic Design from Simulation Traces: A Case Study in Dynamic Production and AGV Scheduling

Model ReleasesDGX agent

arXiv:2608.09343v1 Announce Type: new Abstract: Simulation-based optimization (SBO) evaluates executable policies under stochastic dynamics, but most methods treat the simulator as a black box: aggreg

M^3Prune: Hierarchical Communication Graph Pruning for Efficient Multi-Modal Multi-Agent Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2511.19969v2 Announce Type: replace Abstract: Recent advancements in multi-modal retrieval-augmented generation (mRAG), which enhance multi-modal large language models (MLLMs) with external know

'Many Are My Names': The Anatomy of the Assistant and Its Personas via Sparse Autoencoders

ResearchDGX agent

arXiv:2608.07852v1 Announce Type: new Abstract: How a language model internally represents who is speaking, the Assistant, an assigned roleplay persona, or a narrated story character, remains underexp

MELLON - Multimodal Enhanced LLM for Online Navigation

Model ReleasesDGX agent

arXiv:2608.09121v1 Announce Type: new Abstract: Web navigation agents are capable of addressing various types of tasks on different websites. Current baselines on web navigation are either unimodal or

Memory-Efficient Activation Checkpointing with Sliding Window and Hirschberg's Algorithm for 0/1 Knapsack Solving in PyTorch

Model ReleasesDGX agent

arXiv:2608.08740v1 Announce Type: new Abstract: Activation checkpointing minimizes the runtime of neural networks under a given memory budget, by selecting which intermediate tensors to store and whic

Mesh-Attention: A New Communication-Efficient Distributed Attention with Improved Data Locality

HardwareDGX agent

arXiv:2512.20968v2 Announce Type: replace-cross Abstract: Distributed attention is essential for scaling large language models (LLMs) to long contexts, yet existing methods either have limited paralle

Metadata Reconstruction from Values Alone: Recovering Column Semantics in Undocumented Warehouses

ApplicationsDGX agent

arXiv:2608.07946v1 Announce Type: cross Abstract: Text-to-SQL benchmarks ship schemas whose column names already say what the columns mean. Production warehouses are the inverse: cryptic identifiers,

MetaSpace: Metamorphic Testing for Spatial Cognition in Embodied Agents

Model ReleasesDGX agent

arXiv:2608.07533v1 Announce Type: new Abstract: An embodied agent is an intelligent entity that interacts with its environment through a physical body. Currently, the evaluation of embodied agents pri

Mind the Hook: Source-Level Auditing of Privacy Defenses in Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2608.09001v1 Announce Type: cross Abstract: Black-box privacy scores for retrieval-augmented generation (RAG) are difficult to interpret unless the audited defense's active pipeline hook is know

MPISuperRes-PnP: A Super-Resolution Zero-Shot Plug-and-Play Reconstruction Algorithm for Magnetic Particle Imaging

Model ReleasesDGX agent

arXiv:2608.09672v1 Announce Type: new Abstract: Magnetic Particle Imaging (MPI) is an emerging medical imaging modality. MPI is based on the non-linear response of magnetic nanoparticles to an applied

Multimodal Federated Learning under Dual-Axis Modality Missingness

Local AiDGX agent

arXiv:2608.09240v1 Announce Type: cross Abstract: Multimodal federated learning (FL) supports collaborative modeling in privacy-sensitive health-sensing and medical settings, but realistic deployments

NBA_Streaming: A Large-Scale Benchmark for Fine-Grained Basketball Commentary Generation in Continuous Streams

Model ReleasesDGX agent

arXiv:2608.09200v1 Announce Type: new Abstract: Live basketball commentary generation requires determining when an event is sufficiently observable and describing it before subsequent events unfold. H

NeuroGuard: Neural Gradient Update Aware of Representation Damage

Model ReleasesDGX agent

arXiv:2608.08068v1 Announce Type: new Abstract: Long-tailed class-incremental learning (LT-CIL) must learn new classes from imbalanced streams while retaining old classes. Existing methods mainly chan

NewtonGS: Physics-Structured Object-Level Neural Newtonian Dynamics for Gaussian Scene Animation

ResearchDGX agent

arXiv:2608.07598v1 Announce Type: new Abstract: Animating objects in a static 3D Gaussian scene requires an explicit object-level dynamic state and a controllable model of object motion. Existing dyna

No Unique Minimizer, No Problem: On the Consistency of Robust Neural Classifiers

Model ReleasesDGX agent

arXiv:2608.08489v1 Announce Type: new Abstract: Neural network classifiers trained by cross-entropy minimization are highly sensitive to label noise and adversarial contamination. While robust alterna

Now in preview: The ChatGPT desktop app for Linux. Use ChatGPT, ChatGPT Work, and Codex where you already work and build, with your projects…

Model ReleasesDGX agent

OpenAI released a preview of its new ChatGPT desktop application for Linux on August 11 2026. The app allows users to run ChatGPT, ChatGPT Work and Codex directly from their desktop, integrating with

Open-World Hierarchical Perception: Taxonomic Abstraction over Class-Agnostic Proposals for the Safe Handling of Out-of-Vocabulary Road Objects

Model ReleasesDGX agent

arXiv:2608.07577v1 Announce Type: cross Abstract: A closed-set detector for autonomous driving must assign every object one of a fixed set of labels. On an object outside that set (a horse-drawn carri

OpenAI launches a ChatGPT desktop app for Linux in preview, supporting ChatGPT Work, Codex, and more; Computer Use outside the in-app browser is not available (Frederic Lardinois/The New Stack)

Model ReleasesDGX agent

Frederic Lardinois / The New Stack: OpenAI launches a ChatGPT desktop app for Linux in preview, supporting ChatGPT Work, Codex, and more; Computer Use outside the in-app browser is not available — On

Optimal Multi-Agent Path Finding in Continuous Time

Model ReleasesDGX agent

arXiv:2508.16410v3 Announce Type: replace-cross Abstract: Continuous-time Conflict Based Search (CCBS) has been widely used as an exact baseline for Continuous-time Multi-Agent Path Finding (MAPFR), a

Parameter-Dependent LMI Synthesis for Semi-Global Differential ISS Trajectory Tracking of Nonholonomic Mobile Robots Under Multiplicative Wheel Slip

Model ReleasesDGX agent

arXiv:2608.08049v1 Announce Type: cross Abstract: This paper presents a parameter-dependent linear matrix inequality (LMI) framework for trajectory tracking of nonholonomic mobile robots subject to se

Parameter Exploration for RLVR via Variational Learning

Model ReleasesDGX agent

arXiv:2608.09805v1 Announce Type: cross Abstract: Exploration has been a focus of reinforcement learning research for a long time. Recently, there has been growing evidence that it is also an importan

Path-dependent Discrete Amortized Inference

Model ReleasesDGX agent

arXiv:2608.08644v1 Announce Type: new Abstract: We consider the problem of sampling compositional and discrete objects from a given unnormalized posterior distribution. Notably, recent studies have sh

Personalized Federated Learning via Variance-Aware Nonparametric Empirical Bayes

Model ReleasesDGX agent

arXiv:2608.09074v1 Announce Type: cross Abstract: We develop a new approach to Personalized Federated Learning across heterogeneous clients using Nonparametric Empirical Bayes (NPEB). Leveraging the a

Personalized Lower-limb Exoskeleton Assistance via Preference-based Bayesian Optimization

Model ReleasesDGX agent

arXiv:2608.09015v1 Announce Type: new Abstract: A significant challenge in exoskeleton robotics is the need to dynamically adapt control profiles to individual motion preferences, thereby ensuring bot

Predictive safety filter enhanced curriculum learning control for efficient vehicle dynamics controller

Model ReleasesDGX agent

arXiv:2608.09653v1 Announce Type: cross Abstract: Recent advances in learning-based control have enabled impressive achievements in solving complex control problems in various domains. However, since

Private Etymology: Designing Relational Reuse of Shared Symbols in Long-Term Human-AI Interaction

Local AiDGX agent

arXiv:2608.08443v1 Announce Type: cross Abstract: Previous studies have shown that people can develop shared symbols, partner-specific expressions, personal idioms, inside jokes, and other parts of a

Probabilistic Circuits for Knowledge Graph Completion with Reduced Rule Sets

Model ReleasesDGX agent

arXiv:2508.06706v2 Announce Type: replace Abstract: Rule-based methods for knowledge graph completion provide explainable results, but often require tens of thousands of rules to achieve competitive p

Psychological methods really work. Has anyone tried to encouraging and instilling confidence to GPT?

Model ReleasesDGX agent

Oddly enough, encouragement actually influences the performance of not only Claude but also GPT and other AIs. While Claude was working on a complex problem related to the Riemann Hypothesis, the Anth

Query-Only Backdoor Attacks on Self-Evolving Skills via Trajectory Poisoning

AgentsDGX agent

arXiv:2608.08303v1 Announce Type: new Abstract: Agentic skills improve large language model (LLM) agents by encoding reusable procedures for complex tasks. However, manually authored skills often adap

RAG-Based Auto-Configuration for Industrial Fieldbus Devices

Model ReleasesDGX agent

arXiv:2608.08618v1 Announce Type: cross Abstract: Industrial device commissioning requires engineers to manually extract hundreds of protocol-specific parameters from heterogeneous PDF manuals and tra

Reading is not Reasoning: Bridging the Agentic Policy Gap in Vision-Text Compression

SafetyDGX agent

arXiv:2608.08960v1 Announce Type: new Abstract: Multi-step language-model agents repeatedly process growing interaction histories, leading to substantial context costs. Vision--text compression reduce

RecoverFly: A Failure-Aware Reinforcement Learning Post-Training Framework for Aerial Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2608.09467v1 Announce Type: cross Abstract: Unmanned aerial vehicle vision-language navigation (UAV-VLN) requires agents to translate visual observations and language instructions into reliable

Reflex First, Reflect Later: Latency-Aware Embodied LLM Agents for Dynamic Response

SafetyDGX agent

arXiv:2506.07223v2 Announce Type: replace Abstract: Large language models (LLMs) have substantially improved the planning capabilities of embodied agents, enabling their deployment in dynamic and safe

RenderMatte: Exact-Alpha Rendering and Group-Relative Alignment for Image Matting

Model ReleasesDGX agent

arXiv:2608.08487v1 Announce Type: new Abstract: Image matting is an essential enabling technology for modern visual content production, where foreground extraction determines the realism and editabili

Rethinking 3D Segmentation from Individual LiDAR Scans: Incidence-Aware Sampling on the SIP Benchmark

Model ReleasesDGX agent

arXiv:2608.07757v1 Announce Type: new Abstract: 3D scene understanding is increasingly important in construction, yet most methods are developed on curated datasets that do not fully reflect real site

Rethinking Attention Locality in Spiking Transformers

Model ReleasesDGX agent

arXiv:2608.08541v1 Announce Type: new Abstract: Spiking Transformers provide a promising paradigm for efficient visual processing with spike-driven computation, yet their Softmax-free Spiking Self-Att

Rethinking Factor Sharing in Federated LoRA: A Rank-Aware Adaptive Approach

Local AiDGX agent

arXiv:2608.09742v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) represents large language model (LLM) updates with two compact matrix factors, i.e., A and B, providing an efficient way to

SAGE: SLO-Aware Adaptive Retrieval for Production RAG Systems

Model ReleasesDGX agent

arXiv:2608.08237v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems in production operate under strict service level objectives (SLOs) on tail latency and infrastructure cos

SAIN: Structure-Aware Interactive Navigation with Active Dialogue Grounding for Mobile Robot

Model ReleasesDGX agent

arXiv:2608.09196v1 Announce Type: new Abstract: Most existing vision-language navigation tasks assume that instructions are complete and unambiguous. However, real-world robots often encounter natural

SC-Diff: Semantically Calibrated Diffusion for Visible-to-Infrared Image Translation

SafetyDGX agent

arXiv:2608.08555v1 Announce Type: new Abstract: Visible-to-infrared image translation provides a practical way to expand infrared training data using abundant visible images. Diffusion models are prom

SeqLoc: Beyond the Single Frame for Cross-View Geo-Localization in Feature-Sparse Scenes

Model ReleasesDGX agent

arXiv:2608.07835v1 Announce Type: new Abstract: Cross-View Geo-Localization (CVGL) with OpenStreetMap (OSM) performs well in structure-rich urban environments but collapses in feature-sparse scenes su

Signature-Guided Capacity Occupancy for Dense Expert Merging

TutorialsDGX agent

arXiv:2608.09201v1 Announce Type: new Abstract: Dense expert merging combines domain-specialized language models into one single checkpoint, typically by admitting task-vector support in weight space.

SNR-Edit: Structure-Aware Noise Rectification for Inversion-Free Flow-Based Editing

ResearchDGX agent

arXiv:2601.19180v2 Announce Type: replace-cross Abstract: Inversion-free image editing using flow-based generative models challenges the prevailing inversion-based pipelines. However, existing approac

Spectral Outliers Reveal Dominant Learned Structure in Transformer Attention

Model ReleasesDGX agent

arXiv:2608.07921v1 Announce Type: cross Abstract: We apply Marchenko-Pastur (MP) random matrix theory to pre-trained attention weights in order to separate each projection matrix into a random-like bu

Structured Phonological Representations for Audio-Articulatory rtMRI Speech Classification

ResearchDGX agent

arXiv:2608.09767v1 Announce Type: new Abstract: Real-time MRI makes it possible to observe vocal-tract articulation during speech, but mapping these articulatory patterns to phonetic and phonological

SUM-AgriVLN: Spatial Understanding Memory for Agricultural Vision-and-Language Navigation

Model ReleasesDGX agent

arXiv:2510.14357v2 Announce Type: replace-cross Abstract: Agricultural robots are emerging as powerful assistants across a wide range of agricultural tasks, nevertheless, they are still heavily relyin

Task-Oriented Formation Decision via Reinforcement Learning: Herding an Attacking Swarm

Model ReleasesDGX agent

arXiv:2608.09258v1 Announce Type: new Abstract: Multi-robot systems can accomplish tasks that are difficult for a single robot by organizing into task-specific formations. Different from existing stud

← Previous
1…602603604605606…1060
Next →