AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
19 May 2026

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction

Model ReleasesDGX agent

arXiv:2605.17360v1 Announce Type: new Abstract: Real-time duplex interaction is essential for multimodal AI systems operating in real-world scenarios, where models must continuously process streaming

OmniCode: A Benchmark for Evaluating Software Engineering Agents

Model ReleasesDGX agent

arXiv:2602.02262v3 Announce Type: replace-cross Abstract: LLM-powered coding agents are redefining how real-world software is developed. To drive the research towards better coding agents, we require

OmniPro: A Comprehensive Benchmark for Omni-Proactive Streaming Video Understanding

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.18577v1 Announce Type: new Abstract: Omni-proactive streaming video understanding, i.e., autonomously deciding when to speak and what to say from continuous audio-visual streams, is an emer

OmniVL-Guard Pro: A Tool-Augmented Agent for Omnibus Vision-Language Forensics

Model ReleasesDGX agent

arXiv:2605.16962v1 Announce Type: cross Abstract: Existing vision-language forgery detection and grounding methods operate under a closed-world paradigm, assuming verification can be completed by the

On Improving Multimodal Pedestrian Trajectory Prediction with CVAE: A Study on Benchmark and Robot Data

Model ReleasesDGX agent

arXiv:2605.18262v1 Announce Type: new Abstract: Accurate pedestrian trajectory prediction is crucial for autonomous systems operating in complex environments, such as modular buses and delivery robots

One Hand to Rule Them All: Canonical Representations for Unified Dexterous Manipulation

Model ReleasesDGX agent

arXiv:2602.16712v2 Announce Type: replace Abstract: Dexterous manipulation policies today largely assume fixed hand designs, severely restricting their generalization to new embodiments with varied ki

One Model, Two Roles: Emergent Specialization in a Shared Recurrent Transformer

Model ReleasesDGX agent

arXiv:2605.17811v1 Announce Type: cross Abstract: Can a shared-weight recurrent Transformer develop distinct internal roles without being partitioned into separate modules? We study this in Asymmetric

One of the loudest applauses in the entire Google keynote: Nishtha put on the new Gentle Monster + Gemini glasses, tapped the side to summon…

Model ReleasesDGX agent

One of the loudest applauses in the entire Google keynote: Nishtha put on the new Gentle Monster + Gemini glasses, tapped the side to summon Gemini, and ALL in one prompt said “take a photo and put a

'One thing that we've been seeing recently is that inference benchmarks don't really match production workloads that well.' - @realDanFu, VP…

Model ReleasesDGX agent

'One thing that we've been seeing recently is that inference benchmarks don't really match production workloads that well.' - @realDanFu, VP of Kernels When you're running dozens of concurrent coding

Online Resource Allocation with Convex-set Machine-Learned Advice

Model ReleasesDGX agent

arXiv:2306.12282v2 Announce Type: replace-cross Abstract: Decision-makers often have access to machine-learned predictions about future demand that can help guide online resource allocation decisions.

open sourcing Marlin-2B 🐟 a tiny VLM to extract structured information from videos Marlin is finetuned for two questions devs want to ask i…

Model ReleasesDGX agent

open sourcing Marlin-2B 🐟 a tiny VLM to extract structured information from videos Marlin is finetuned for two questions devs want to ask in their videos: what is happening, and when? Best open model

OpenJarvis: Personal AI, On Personal Devices

Model ReleasesDGX agent

arXiv:2605.17172v1 Announce Type: cross Abstract: Personal AI stacks, like OpenClaw and Hermes Agent, are becoming central to daily work, yet they route nearly every query (often over sensitive local

Optimising CSRNet with parameter-free attention mechanisms for crowd counting in public transport

Model ReleasesDGX agent

arXiv:2605.18349v1 Announce Type: cross Abstract: Occupancy estimation and crowd counting are critical tasks in designing smart and efficient public transport vehicles. Given that public transport loa

ORACLE: Anticipating Scams from Partial Trajectories in Streaming App Usage

Model ReleasesDGX agent

arXiv:2605.16363v1 Announce Type: new Abstract: Smartphone scams are increasingly prevalent and typically manifest as multi-stage, cross-application processes with gradually emerging intent. Effective

Ordinal Adaptive Correction: A Data-Centric Approach to Ordinal Image Classification with Noisy Labels

Model ReleasesDGX agent

arXiv:2509.02351v3 Announce Type: replace-cross Abstract: Labeled data is a fundamental component in training supervised deep learning models for computer vision tasks. However, the labeling process,

OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents

Model ReleasesDGX agent

arXiv:2506.16042v2 Announce Type: replace Abstract: Generative AI is being leveraged to solve a variety of computer-use tasks involving desktop applications. State-of-the-art systems have focused sole

Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks

Model ReleasesDGX agent

arXiv:2605.18583v1 Announce Type: cross Abstract: Coding agents now run autonomously with shell, file, and network privileges. When a user issues a benign request, the agent sometimes does more than a

PACE: Geometry-Aware Bridge Transport for Single-Cell Trajectory Inference

Model ReleasesDGX agent

arXiv:2605.18587v1 Announce Type: cross Abstract: Single-cell trajectory inference from destructive time-course snapshots is fundamentally ill-posed: neither cross-time cell correspondences nor contin

PaliBench: A Multi-Reference Blueprint for Classical Language Translation Benchmarks

Model ReleasesDGX agent

arXiv:2605.16881v1 Announce Type: new Abstract: Digital humanities projects increasingly rely on machine translation and large language models to widen access to classical, religious, and otherwise un

PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts

Model ReleasesDGX agent

arXiv:2605.17028v1 Announce Type: cross Abstract: Large language models (LLMs) hallucinate with confidence: their outputs can be fluent, authoritative, and simply wrong. In medical, legal, and scienti

Parameter-Efficient Domain Adaptation of Physics-Informed Self-Attention based GNNs for AC Power Flow Prediction

Model ReleasesDGX agent

arXiv:2602.18227v2 Announce Type: replace Abstract: Accurate AC power flow (AC-PF) prediction under domain shift is critical when models trained on medium-voltage (MV) grids are deployed on high-volta

PAREDA: A Multi-Accent Speech Dataset of Natural Language Processing Research Discussions

Model ReleasesDGX agent

arXiv:2605.17860v1 Announce Type: cross Abstract: While modern Automatic Speech Recognition (ASR) systems achieve high accuracy on benchmark corpora, their performance often degrades when there is rea

PEGRL: Improving Machine Translation by Post-Editing Guided Reinforcement Learning

Model ReleasesDGX agent

arXiv:2602.03352v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown strong promise for LLM-based machine translation, with recent methods such as GRPO demonstrating notable gains

People are generating over 1.5 billion images a week in ChatGPT. Researcher @kenjihata joins Product lead @adele__li and host @AndrewMayne t…

Model ReleasesDGX agent

People are generating over 1.5 billion images a week in ChatGPT. Researcher @kenjihata joins Product lead @adele__li and host @AndrewMayne to explore the new use cases and trends emerging since the la

Perfect Parallelization in Mini-Batch SGD with Classical Momentum Acceleration

Model ReleasesDGX agent

arXiv:2605.18609v1 Announce Type: new Abstract: Accelerating stochastic gradient methods with classical momentum schemes, such as Polyak's heavy ball, has proven highly successful in training large-sc

PERL: Parameter Efficient Reasoning in CLIP Latent Space

Model ReleasesDGX agent

arXiv:2605.18464v1 Announce Type: new Abstract: Contrastively trained vision-language models such as CLIP provide strong zero-shot transfer by aligning images and text in a shared embedding space. How

PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and Realistic Task Environments

Model ReleasesDGX agent

arXiv:2603.23231v2 Announce Type: replace Abstract: Empowering large language models with long-term memory is crucial for building agents that adapt to users' evolving needs. Existing evaluations of t

PESD-TSF: A Period-Aware and Explicit Structured Decomposition Framework for Long-Term Time Series Forecasting

Model ReleasesDGX agent

arXiv:2605.16449v1 Announce Type: cross Abstract: Deep forecasting models often suffer from attenuated periodic perception and entangled trend-noise representations as network depth increases. Moreove

PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media

Model ReleasesDGX agent

arXiv:2605.17187v1 Announce Type: cross Abstract: Social media are shifting towards pluralism -- community-governed platforms where groups define their own norms. What violates rules in one community

Position: AI Evaluations Should be Grounded on a Theory of Capability

Model ReleasesDGX agent

arXiv:2509.19590v2 Announce Type: replace Abstract: Evaluations of generative models are now ubiquitous, and their outcomes critically shape public and scientific expectations of AI's capabilities. Ye

POST: Prior-Observation Adversarial Learning of Spatio-Temporal Associations for Multivariate Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.18128v1 Announce Type: new Abstract: Existing Multivariate Time Series Anomaly Detection (MTSAD) frameworks increasingly rely on integrating Graph Neural Networks (GNNs) with sequence model

Post-Trained MoE Can Skip Half Experts via Self-Distillation

Model ReleasesDGX agent

arXiv:2605.18643v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) scales language models efficiently through sparse expert activation, and its dynamic variant further reduces computation by a

Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency

Model ReleasesDGX agent

arXiv:2605.18732v1 Announce Type: cross Abstract: While scaling laws govern aggregate large language model performance, no scaling law has linked factual recall to both model size and training-data co

PriHA: A RAG-Enhanced LLM Framework for Primary Healthcare Assistant in Hong Kong

Model ReleasesDGX agent

arXiv:2604.14215v2 Announce Type: replace-cross Abstract: To address the unsustainable rise in public health expenditures, the Hong Kong SAR Government is shifting its strategic focus to primary healt

PRIME: Physically-consistent Robotic Inertial and Motion Estimation for Legged and Humanoid Robots

Model ReleasesDGX agent

arXiv:2605.17681v1 Announce Type: new Abstract: Humanoid and legged robots interact with the environment through intermittent contacts, making accurate motion estimation fundamentally dependent on rea

PRISMat: Policy-Driven, Permutation-Invariant Autoregressive Material Generation

Model ReleasesDGX agent

arXiv:2605.16612v1 Announce Type: new Abstract: Rapid identification of candidate materials with target properties has become a key task in materials science. Machine learning has emerged as an altern

Privacy Policy Enforcement Guardrails for Data-Sensitive Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2605.17034v1 Announce Type: cross Abstract: Standard PII filters often miss contextual data leakage in RAG systems, such as non-regulated attribute clusters that collectively identify individual

Privacy Preserving Reinforcement Learning with One-Sided Feedback

Model ReleasesDGX agent

arXiv:2605.18246v1 Announce Type: cross Abstract: We study reinforcement learning (RL) in multi-dimensional continuous state and action spaces with one-sided feedback, where the agent receives partial

Probing for Representation Manifolds in Superposition

Model ReleasesDGX agent

arXiv:2605.18537v1 Announce Type: cross Abstract: This paper introduces the Manifold Probe, a supervised method for discovering representation manifolds in superposition. The method generalizes linear

Probing SMEFT Operators through tar{t}tar{t} Production with Hyper-Graph Neural Networks at the LHC

Model ReleasesDGX agent

arXiv:2605.18382v1 Announce Type: cross Abstract: We present a phenomenological study of tar{t}tar{t} production in proton-proton collisions at sqrt{s} = 13~TeV, using a Hyper-Graph Neural Network (H-

ProfBench: Multi-Domain Rubrics requiring Professional Knowledge to Answer and Judge

Model ReleasesDGX agent

arXiv:2510.18941v2 Announce Type: replace-cross Abstract: Evaluating progress in large language models (LLMs) is often constrained by the challenge of verifying responses, limiting assessments to task

Prompt Compression in Diffusion Large Language Models: Evaluating LLMLingua-2 on LLaDA

Model ReleasesDGX agent

arXiv:2605.17932v1 Announce Type: cross Abstract: Prompt compression reduces inference cost and context length in large language models, but prior evaluations focus primarily on autoregressive archite

Prompt2Fingerprint: Plug-and-Play LLM Fingerprinting via Text-to-Weight Generation

Model ReleasesDGX agent

arXiv:2605.18474v1 Announce Type: cross Abstract: The widespread deployment and redistribution of large language models (LLMs) have made model provenance tracking a critical challenge. While existing

Prompts Don't Protect: Architectural Enforcement via MCP Proxy for LLM Tool Access Control

Model ReleasesDGX agent

arXiv:2605.18414v1 Announce Type: cross Abstract: Large language models increasingly operate as autonomous agents that select and invoke tools from large registries. We identify a critical gap: when u

Protection Is (Nearly) All You Need: Structural Protection Dominates Scoring in Globally Capped KV Eviction

Model ReleasesDGX agent

arXiv:2605.18053v1 Announce Type: cross Abstract: We study KV cache eviction under a shared globally capped decode-time harness. Seven policies (LRU, H2O, SnapKV, StreamingLLM, Ada-KV, QUEST, Random)

Protein Fold Classification at Scale: Benchmarking and Pretraining

Model ReleasesDGX agent

arXiv:2605.18552v1 Announce Type: new Abstract: Classifying protein topology is essential for deciphering biological function, but progress is held back by the lack of large-scale benchmarks that avoi

ProtoSiTex: Learning Semi-Interpretable Prototypes for Multi-label Text Classification

Model ReleasesDGX agent

arXiv:2510.12534v4 Announce Type: replace Abstract: The rapid growth of user-generated text across digital platforms has intensified the need for interpretable models capable of fine-grained text clas

Provable Knowledge Acquisition and Extraction in One-Layer Transformers

Model ReleasesDGX agent

arXiv:2508.00901v4 Announce Type: replace-cross Abstract: Large language models may encounter factual knowledge during pre-training yet fail to reliably use that knowledge after fine-tuning. Despite g

Provably Shorter Scratchpads in Hybrid DeltaNet-Attention Decoders

Model ReleasesDGX agent

arXiv:2605.16640v1 Announce Type: new Abstract: We investigate the expressive power of hybrid recurrent-attention decoders, a class of architectures used in recent open-source language models such as

ProxyKV: Cross-Model Proxy Pruning for Efficient Long-Context LLM Inference

Model ReleasesDGX agent

arXiv:2605.16360v1 Announce Type: cross Abstract: Efficient long-context inference in Large Language Models (LLMs) is severely constrained by the Key-Value (KV) cache memory wall, yet existing pruning

QLIF-CAST: Quantum Leaky-Integrate-and-Fire for Time-Series Weather Forecasting

Model ReleasesDGX agent

arXiv:2605.18333v1 Announce Type: cross Abstract: Accurate and efficient time-series forecasting remains a challenging problem for both classical and quantum neural architectures, particularly in mult

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi

Model ReleasesDGX agent

arXiv:2605.18380v1 Announce Type: new Abstract: We introduce an extensive qualitative spatial and temporal reasoning (QSTR) benchmark for evaluating large language models (LLMs). We pose questions con

QuCo-RAG: Quantifying Uncertainty from the Pre-training Corpus for Dynamic Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2512.19134v2 Announce Type: replace Abstract: Dynamic Retrieval-Augmented Generation adaptively determines when to retrieve during generation to mitigate hallucinations in large language models

Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models

Model ReleasesDGX agent

arXiv:2501.17549v2 Announce Type: replace Abstract: Graph-structured data plays a vital role in numerous domains, such as social networks, citation networks, commonsense reasoning graphs and knowledge

Queue Length Regret Bounds for Contextual Queueing Bandits

Model ReleasesDGX agent

arXiv:2601.19300v2 Announce Type: replace Abstract: We introduce contextual queueing bandits, a new context-aware framework for scheduling while simultaneously learning unknown service rates. Individu

Radial-Angular Geometry for Reliable Update Diagnosis in Noisy-Label Learning

Model ReleasesDGX agent

arXiv:2605.17429v1 Announce Type: cross Abstract: Noisy-label methods often estimate sample reliability from forward-space signals such as loss, confidence, or entropy. These signals indicate whether

RAP: Runtime Adaptive Pruning for LLM Inference

Model ReleasesDGX agent

arXiv:2505.17138v5 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at language understanding and generation, but their enormous computational and memory requirements hinder d

ReBaR: Reference-Based Reasoning for Robust Pose Estimation from Monocular Images

Model ReleasesDGX agent

arXiv:2303.11675v3 Announce Type: replace Abstract: R}easoning for Robust Human Pose and Shape Estimation), designed to estimate human body shape and pose from single-view images. ReBaR effectively ad

REBAR: Reference Ethical Benchmark for Autonomy Readiness

Model ReleasesDGX agent

arXiv:2605.18423v1 Announce Type: new Abstract: As autonomous systems grow more advanced, objective metrics to evaluate their ethical and legal compliance are critical for informing end users of their

Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts

Model ReleasesDGX agent

arXiv:2510.07239v2 Announce Type: replace Abstract: Automated red-teaming has emerged as a scalable approach for auditing Large Language Models (LLMs) prior to deployment, yet existing approaches lack

← Previous
1…237238239240241…377
Next →