AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,611 results
Model Releases

Probing for Representation Manifolds in Superposition

DGX agent

arXiv:2605.18537v1 Announce Type: cross Abstract: This paper introduces the Manifold Probe, a supervised method for discovering representation manifolds in superposition. The method generalizes linear

model-releasesarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Probing SMEFT Operators through tar{t}tar{t} Production with Hyper-Graph Neural Networks at the LHC

DGX agent

arXiv:2605.18382v1 Announce Type: cross Abstract: We present a phenomenological study of tar{t}tar{t} production in proton-proton collisions at sqrt{s} = 13~TeV, using a Hyper-Graph Neural Network (H-

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ProfBench: Multi-Domain Rubrics requiring Professional Knowledge to Answer and Judge

DGX agent

arXiv:2510.18941v2 Announce Type: replace-cross Abstract: Evaluating progress in large language models (LLMs) is often constrained by the challenge of verifying responses, limiting assessments to task

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Prompt Compression in Diffusion Large Language Models: Evaluating LLMLingua-2 on LLaDA

DGX agent

arXiv:2605.17932v1 Announce Type: cross Abstract: Prompt compression reduces inference cost and context length in large language models, but prior evaluations focus primarily on autoregressive archite

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Prompt2Fingerprint: Plug-and-Play LLM Fingerprinting via Text-to-Weight Generation

DGX agent

arXiv:2605.18474v1 Announce Type: cross Abstract: The widespread deployment and redistribution of large language models (LLMs) have made model provenance tracking a critical challenge. While existing

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Prompts Don't Protect: Architectural Enforcement via MCP Proxy for LLM Tool Access Control

DGX agent

arXiv:2605.18414v1 Announce Type: cross Abstract: Large language models increasingly operate as autonomous agents that select and invoke tools from large registries. We identify a critical gap: when u

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Protection Is (Nearly) All You Need: Structural Protection Dominates Scoring in Globally Capped KV Eviction

DGX agent

arXiv:2605.18053v1 Announce Type: cross Abstract: We study KV cache eviction under a shared globally capped decode-time harness. Seven policies (LRU, H2O, SnapKV, StreamingLLM, Ada-KV, QUEST, Random)

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Protein Fold Classification at Scale: Benchmarking and Pretraining

DGX agent

arXiv:2605.18552v1 Announce Type: new Abstract: Classifying protein topology is essential for deciphering biological function, but progress is held back by the lack of large-scale benchmarks that avoi

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

ProtoSiTex: Learning Semi-Interpretable Prototypes for Multi-label Text Classification

DGX agent

arXiv:2510.12534v4 Announce Type: replace Abstract: The rapid growth of user-generated text across digital platforms has intensified the need for interpretable models capable of fine-grained text clas

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Provable Knowledge Acquisition and Extraction in One-Layer Transformers

DGX agent

arXiv:2508.00901v4 Announce Type: replace-cross Abstract: Large language models may encounter factual knowledge during pre-training yet fail to reliably use that knowledge after fine-tuning. Despite g

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Provably Shorter Scratchpads in Hybrid DeltaNet-Attention Decoders

DGX agent

arXiv:2605.16640v1 Announce Type: new Abstract: We investigate the expressive power of hybrid recurrent-attention decoders, a class of architectures used in recent open-source language models such as

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

ProxyKV: Cross-Model Proxy Pruning for Efficient Long-Context LLM Inference

DGX agent

arXiv:2605.16360v1 Announce Type: cross Abstract: Efficient long-context inference in Large Language Models (LLMs) is severely constrained by the Key-Value (KV) cache memory wall, yet existing pruning

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

QLIF-CAST: Quantum Leaky-Integrate-and-Fire for Time-Series Weather Forecasting

DGX agent

arXiv:2605.18333v1 Announce Type: cross Abstract: Accurate and efficient time-series forecasting remains a challenging problem for both classical and quantum neural architectures, particularly in mult

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi

DGX agent

arXiv:2605.18380v1 Announce Type: new Abstract: We introduce an extensive qualitative spatial and temporal reasoning (QSTR) benchmark for evaluating large language models (LLMs). We pose questions con

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

QuCo-RAG: Quantifying Uncertainty from the Pre-training Corpus for Dynamic Retrieval-Augmented Generation

DGX agent

arXiv:2512.19134v2 Announce Type: replace Abstract: Dynamic Retrieval-Augmented Generation adaptively determines when to retrieve during generation to mitigate hallucinations in large language models

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models

DGX agent

arXiv:2501.17549v2 Announce Type: replace Abstract: Graph-structured data plays a vital role in numerous domains, such as social networks, citation networks, commonsense reasoning graphs and knowledge

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Queue Length Regret Bounds for Contextual Queueing Bandits

DGX agent

arXiv:2601.19300v2 Announce Type: replace Abstract: We introduce contextual queueing bandits, a new context-aware framework for scheduling while simultaneously learning unknown service rates. Individu

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Radial-Angular Geometry for Reliable Update Diagnosis in Noisy-Label Learning

DGX agent

arXiv:2605.17429v1 Announce Type: cross Abstract: Noisy-label methods often estimate sample reliability from forward-space signals such as loss, confidence, or entropy. These signals indicate whether

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

RAP: Runtime Adaptive Pruning for LLM Inference

DGX agent

arXiv:2505.17138v5 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at language understanding and generation, but their enormous computational and memory requirements hinder d

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ReBaR: Reference-Based Reasoning for Robust Pose Estimation from Monocular Images

DGX agent

arXiv:2303.11675v3 Announce Type: replace Abstract: R}easoning for Robust Human Pose and Shape Estimation), designed to estimate human body shape and pose from single-view images. ReBaR effectively ad

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

REBAR: Reference Ethical Benchmark for Autonomy Readiness

DGX agent

arXiv:2605.18423v1 Announce Type: new Abstract: As autonomous systems grow more advanced, objective metrics to evaluate their ethical and legal compliance are critical for informing end users of their

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts

DGX agent

arXiv:2510.07239v2 Announce Type: replace Abstract: Automated red-teaming has emerged as a scalable approach for auditing Large Language Models (LLMs) prior to deployment, yet existing approaches lack

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Reference anything: Gemini Omni extends Gemini's native multimodality, allowing you to blend combinations of text, audio, image, and video i…

DGX agent

Gemini Omni is an extension of Google's Gemini model that enhances its multimodal capabilities by enabling seamless integration of text, audio, image, and video inputs and outputs. This advancement al

model-releasesgoogle-ai--x
19 May 2026
Model Releases

Residual Semantic Decomposition of Word Embeddings

DGX agent

arXiv:2605.17482v1 Announce Type: new Abstract: We introduce Residual Semantic Decomposition (RSD), a neural additive decomposition of word embeddings that balances embedding reconstruction with relat

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Responsible Agentic AI Requires Explicit Provenance

DGX agent

arXiv:2605.17169v1 Announce Type: new Abstract: Agentic AI is rapidly proliferating across diverse real-world domains such as software engineering, yet public trust has not kept pace. The central reas

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Rethinking GNNs and Missing Features: Challenges, Evaluation and a Robust Solution

DGX agent

arXiv:2601.04855v2 Announce Type: replace-cross Abstract: Handling missing node features is a key challenge for deploying Graph Neural Networks (GNNs) in real-world domains such as healthcare and sens

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free

DGX agent

arXiv:2605.16767v1 Announce Type: new Abstract: Multi-label legal annotation requires assigning multiple labels from large, evolving taxonomies to long, fact-intensive documents, often under limited s

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Reverse-Engineering Model Editing on Language Models

DGX agent

arXiv:2602.10134v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are pretrained on corpora containing trillions of tokens and, therefore, inevitably memorize sensitive informatio

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Revisiting the Adam-SGD Gap in LLM Pre-Training: The Role of Large Effective Learning Rates

DGX agent

arXiv:2605.17787v1 Announce Type: new Abstract: It is widely believed that stochastic gradient descent (SGD) performs significantly worse than adaptive optimizers such as Adam in pre-training Large La

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

RIE-Greedy: Regularization-Induced Exploration for Contextual Bandits

DGX agent

arXiv:2603.11276v2 Announce Type: replace-cross Abstract: Real-world contextual bandit problems with complex reward models are often tackled with iteratively trained models, such as boosting trees. Ho

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Ringmaster LMO: Asynchronous Linear Minimization Oracle Momentum Method

DGX agent

arXiv:2605.18174v1 Announce Type: new Abstract: Muon has recently emerged as a strong alternative to AdamW for training neural networks, with encouraging large-scale pretraining results and growing ev

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards

DGX agent

arXiv:2509.21319v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Human Feedback (RLHF) and Reinforcement Learning with Verifiable Rewards (RLVR) are the main RL paradigms used in

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

RoboMME: Benchmarking and Understanding Memory for Robotic Generalist Policies

DGX agent

arXiv:2603.04639v2 Announce Type: replace-cross Abstract: Memory is critical for long-horizon and history-dependent robotic manipulation. Such tasks often involve counting repeated actions or manipula

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ROVR-Open-Dataset: A Large-Scale Depth Dataset for Autonomous Driving

DGX agent

arXiv:2508.13977v3 Announce Type: replace Abstract: Depth estimation is a fundamental component of spatial perception for autonomous driving and other unmanned systems operating in open urban environm

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

rsi is here. jesus https://x.com/nickevanjoseph/status/2056760504949842219?s=46

DGX agent

rsi is here. jesus https://x.com/nickevanjoseph/status/2056760504949842219?s=46 Excited to welcome Andrej to the Pretraining team! He'll be building a team focused on using Claude to accelerate pretra

model-releasesswyx--x
19 May 2026
Model Releases

RTI-Bench: A Structured Dataset for Indian Right-to-Information Decision Analysis

DGX agent

arXiv:2605.16843v1 Announce Type: new Abstract: India's Right to Information Act, 2005 gives every citizen the right to demand information from public authorities, yet in practice most people cannot m

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

SaaSBench: Exploring the Boundaries of Coding Agents in Long-Horizon Enterprise SaaS Engineering

DGX agent

arXiv:2605.17526v1 Announce Type: cross Abstract: As autonomous coding agents become capable of handling increasingly long-horizon tasks, they have gradually demonstrated the potential to complete end

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SafeLens: Deliberate and Efficient Video Guardrails with Fast-and-Slow Screening

DGX agent

arXiv:2605.17610v1 Announce Type: cross Abstract: The rapid growth of online video platforms and AI-generated content has made reliable video guardrails a key challenge for safety and real-world deplo

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

SAM 2++: Tracking Anything at Any Granularity

DGX agent

arXiv:2510.18822v4 Announce Type: replace Abstract: Due to the varying granularity of target states across different tasks, most existing trackers are tailored to a single task, which specificity limi

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

SAME: A Semantically-Aligned Music Autoencoder

DGX agent

arXiv:2605.18613v1 Announce Type: cross Abstract: Latent representations are at the heart of the majority of modern generative models. In the audio domain they are typically produced by a neural-audio

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Scalable Knowledge Editing for Mixture-of-Experts LLMs via Tensor-Structured Updates

DGX agent

arXiv:2605.16686v1 Announce Type: new Abstract: Knowledge editing (KE) provides a lightweight alternative to repeated fine-tuning of LLMs. However, most existing KE methods target dense feed-forward l

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Scale-Equivariant Generative Forecasting: Weight-Tied Dilated Convolutions, Wavelet Scattering Inputs, and Spectral-Consistency Training for Self-Similar Time Series

DGX agent

arXiv:2605.17582v1 Announce Type: new Abstract: Many natural and engineered time series -- equity returns, climate anomalies, turbulent velocities, neural recordings, packet-level network traffic -- a

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Scales++: Compute Efficient Evaluation Subset Selection with Cognitive Scales Embeddings

DGX agent

arXiv:2510.26384v2 Announce Type: replace Abstract: The prohibitive cost of evaluating large language models (LLMs) on comprehensive benchmarks necessitates the creation of small yet representative da

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Scaling Laws for Code: A More Data-Hungry Regime

DGX agent

arXiv:2510.08702v2 Announce Type: replace Abstract: Code Large Language Models (LLMs) are revolutionizing software engineering. However, scaling laws that guide the efficient training are predominantl

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

SCARED-C: Corrected Camera Poses for Endoscopic Depth Estimation

DGX agent

arXiv:2605.16628v1 Announce Type: new Abstract: The SCARED dataset is a widely used benchmark for endoscopic depth estimation, offering ground-truth 3D reconstructions captured with a structured light

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Scheduling That Speaks: An Interpretable Programmatic Reinforcement Learning Framework

DGX agent

arXiv:2605.18454v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) has recently emerged as a promising approach to solve combinatorial optimization problems such as job shop schedulin

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science

DGX agent

arXiv:2605.18630v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as scientific AI as- sistants, and a growing body of benchmarks evaluates their capabilities acro

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SE-GA: Memory-Augmented Self-Evolution for GUI Agents

DGX agent

arXiv:2605.16883v1 Announce Type: new Abstract: Autonomous Graphical User Interface (GUI) agents often struggle with multi-step tasks due to constrained context windows and static policies that fail t

model-releasesarxiv-cs-lg
19 May 2026
← Previous
1…299300301302303…472
Next →