AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Tutorials

Learning to Predict Middle-Layer Attention in MLLMs for Visual Token Prunin

DGX agent

arXiv:2608.06411v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) achieve strong performance across diverse vision-language tasks, but their efficiency is limited by the cost of

tutorialsarxiv-cs-ai
10 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Long-Horizon Agent Trajectory Attribution: A Unified Benchmark and Fine-Grained Annotation Framework

DGX agent

arXiv:2608.06909v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate through long-horizon trajectories involving user instructions, tool use, external observations, a

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Modular TTT: Rethinking Test-Time Training as Composable Modules

DGX agent

arXiv:2608.07110v1 Announce Type: cross Abstract: Test-time training (TTT) views sequence modeling as an online learning problem in which fast weights are updated by an internal learning rule. Despite

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

MultiView-Bench: A Diagnostic Benchmark for World-Centric Multi-View Integration in VLMs

DGX agent

arXiv:2607.08970v2 Announce Type: replace-cross Abstract: Recent benchmarks for VLMs largely assess single- or limited-view perception, leaving untested the core cognitive ability to integrate observa

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

PURe: A Plug-and-Play Product-Unit Residual Module for Vision Networks

DGX agent

arXiv:2505.04397v3 Announce Type: replace-cross Abstract: Modern vision networks are dominated by additive local transformations, whereas explicit multiplicative local interactions remain underexplore

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

Representation-driven Endoscopic Visual Embedding Alignment for Latent Generation

DGX agent

arXiv:2608.07176v1 Announce Type: cross Abstract: Developing foundation generative models for endoscopy is limited by the gap between natural and clinical images and the computational cost of training

safetyarxiv-cs-ai
10 Aug 2026
Safety

Representation Handoffs for OpenArm-Based Laboratory Mobile Manipulation

DGX agent

arXiv:2608.07154v1 Announce Type: cross Abstract: Open-source robotics and foundation models have lowered the barrier to embodied AI, yet language-guided laboratory automation still requires reliable

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

RoRA: Role-Oriented Regional Allocation for Visual Token Pruning in MLLMs

DGX agent

arXiv:2608.07088v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) encode images as long visual token sequences, making prefilling and KV-cache storage expensive. Existing trai

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Simple-OPD: Demystifying Warm-up for On-policy Distillation

DGX agent

arXiv:2608.06802v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts with token-level supervision from teacher models, but its effectiveness can depend str

model-releasesarxiv-cs-cl
10 Aug 2026
Research

Skaling: Chinchilla's Exponents Meet Kaplan's Coupling

DGX agent

arXiv:2608.07222v1 Announce Type: new Abstract: Neural scaling laws are foundational for language model development, yet standard formulations systematically under- and overestimate loss at data-scarc

researcharxiv-cs-cl
10 Aug 2026
Model Releases

Suppress and Diversify: Refining Robust Pathways for Corruption Robustness

DGX agent

arXiv:2608.06712v1 Announce Type: new Abstract: Model robustness against natural image corruptions is essential for safety-critical applications. While existing methods primarily focus on implicit rep

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

YOLO-PEFT: Parameter-Efficient Fine-Tuning on YOLO Family

DGX agent

arXiv:2608.07051v1 Announce Type: new Abstract: Generic parameter-efficient fine-tuning (PEFT) methods transferred from language models can fail silently on real-time detectors, whose heterogeneous op

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

A Paragraph is Worth a Thousand Captions: Rethinking Text Supervision for Vision-Language Retrieval

DGX agent

arXiv:2608.05260v1 Announce Type: new Abstract: Contrastive vision-language models such as CLIP and BLIP are typically trained on short image captions, limiting their ability to retrieve images from d

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New SheetSage-A2S Dataset

DGX agent

arXiv:2608.06165v1 Announce Type: cross Abstract: Existing audio-to-score (A2S) systems primarily focus on classical music, and the application to popular music remains underexplored. This paper first

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Benchmarking and Enhancing LLMs for Rule-Intensive Review of National Standard Documents

DGX agent

arXiv:2608.06312v1 Announce Type: new Abstract: Large language models (LLMs) increasingly support complex professional tasks, yet their capabilities in rule-intensive document review remain insufficie

model-releasesarxiv-cs-cl
7 Aug 2026
Research

Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation

DGX agent

arXiv:2608.05999v1 Announce Type: new Abstract: Vision-language-action (VLA) models have demonstrated remarkable capabilities in robotic manipulation by leveraging pretrained vision-language models. H

researcharxiv-cs-ro
7 Aug 2026
Model Releases

ChronoVision: Temporal Reasoning via Latent State Reconstruction

DGX agent

arXiv:2608.05631v1 Announce Type: new Abstract: Multimodal large language models excel at passive perception but struggle with complex visual cognitive tasks requiring multi-step temporal reasoning. T

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

CNM-BERT: A Drop-In Structural Embedding for Chinese Characters via Ideographic Description Sequences

DGX agent

arXiv:2608.05167v1 Announce Type: new Abstract: Token-based encoders like BERT treat Chinese characters as atomic identifiers, ignoring their recursive orthographic structure. Consequently, models rel

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

DASH: Decoupled Adaptive Surrogate - Acquisition Harness for Automated Bayesian Optimization

DGX agent

arXiv:2608.00641v2 Announce Type: replace Abstract: Bayesian optimization (BO) relies on a surrogate model and an acquisition function, yet the most suitable choices vary across tasks and optimization

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Domain-Grounded Candidate Selection for Agentic Image Editing: A Shadow Removal Case

DGX agent

arXiv:2608.06075v1 Announce Type: cross Abstract: Commercial vision-language models are reshaping computer vision, with visual priors broad enough to rival task-specific systems. This raises a natural

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

DREAM: LLM-based Dynamic Role-playing via Event-Aware Memory Graph

DGX agent

arXiv:2608.05170v1 Announce Type: cross Abstract: Role-playing agents (RPAs) have emerged as a key application of large language models, enabling immersive and high-fidelity character simulation. Accu

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Energy-Guided Flow Matching

DGX agent

arXiv:2608.05811v1 Announce Type: new Abstract: Pixel-space generative models bypass lossy latent compression, yet necessitate joint learning of global structure and fine-grained details in a high-dim

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Evidential Rule Learning for Interpretable Classification with Abstention

DGX agent

arXiv:2608.05859v1 Announce Type: cross Abstract: Interpretable classification often requires more than accurate predictions for real-life deployment: models should be transparent about the evidence b

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?

DGX agent

arXiv:2608.05747v1 Announce Type: new Abstract: Spatial intelligence is fundamental to embodied agents, yet existing benchmarks focus on local spatial perception from single or few viewpoints, overloo

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization

DGX agent

arXiv:2608.06301v1 Announce Type: new Abstract: As LLMs are increasingly deployed within agentic systems, their capabilities depend not only on the model weights but also on the harness: the prompts,

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Hyper-ES: Effective Evolution Strategies for LLM Reasoning via Descent Direction Merging

DGX agent

arXiv:2608.05541v1 Announce Type: new Abstract: Evolution Strategy (ES) is a promising alternative to gradient-based fine-tuning for resource-constrained Large Language Model (LLM) reasoning. However,

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Innocent Panels, Hateful Stories: Evaluating and Detecting Hateful Intent in Multi-Turn Visual Story Generation

DGX agent

arXiv:2608.05210v1 Announce Type: cross Abstract: Picture books and comics have long been used to disseminate hateful narratives because they are easily understood even by children, as exemplified by

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

IPV-Bench: Benchmarking Image Protection Methods under Diverse Image-to-Video Generation Scenarios

DGX agent

arXiv:2603.26154v2 Announce Type: replace Abstract: Image-to-video (I2V) generation models can be misused to animate a single image into a convincing fake video, motivating perturbation-based image pr

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

MoCA: Implicit Social Context Analysis

DGX agent

arXiv:2608.05825v1 Announce Type: new Abstract: Human social communication, such as affection and intent, is often conveyed in highly implicit ways, where underlying meanings are expressed through ind

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

NeSy-RAG: Neuro-Symbolic RAG for Explainable Question Answering

DGX agent

arXiv:2608.06292v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves question answering by grounding large language models (LLMs) in external knowledge such as text corpora. H

model-releasesarxiv-cs-cl
7 Aug 2026
Applications

nnMIL: A generalizable multiple instance learning framework for computational pathology

DGX agent

arXiv:2511.14907v2 Announce Type: replace Abstract: Computational pathology holds substantial promise for improving diagnosis and guiding treatment decisions. Recent pathology foundation models enable

applicationsarxiv-cs-cv
7 Aug 2026
Model Releases

PoolBench: A Benchmark for Pooling Strategies in Concept Representation Evaluation for Decoder-Only LLMs

DGX agent

arXiv:2608.05162v1 Announce Type: new Abstract: Pooling is a consequential but under-examined design choice in decoder-only concept representation work: practitioners must collapse token-level hidden

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Robust Native Language Identification through Agentic Decomposition

DGX agent

arXiv:2509.16666v2 Announce Type: replace Abstract: Large language models (LLMs) often achieve high performance in native language identification (NLI) benchmarks by leveraging superficial contextual

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Schema-Guided Hierarchical Information Extraction and Semantic Evaluation Using Generative AI

DGX agent

arXiv:2608.06167v1 Announce Type: new Abstract: We present a schema-based framework for extracting complex, structured information from unstructured text documents using generative AI, followed by aut

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptation

DGX agent

arXiv:2608.05785v1 Announce Type: cross Abstract: Multilingual text embedding models are commonly adapted using a single training objective across diverse tasks, despite different tasks requiring fund

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Advancing Utility Pole and Sign Detection Through Deep Learning

DGX agent

arXiv:2608.04061v1 Announce Type: new Abstract: Utility poles are an essential part of the infrastructure used to support power distribution systems and other critical public services. Their regular i

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

An active-learning framework for real-time depth perception from monocular vision streams

DGX agent

arXiv:2608.04917v1 Announce Type: new Abstract: Biological visual systems can perceive depth from monocular vision flow, continuously integrating temporal visual cues while maintaining a balance betwe

model-releasesarxiv-cs-cv
6 Aug 2026
Local Ai

Bi-Level Reinforcement Learning Pathway for Sim-to-Real Optimality

DGX agent

arXiv:2510.17709v2 Announce Type: replace-cross Abstract: Training Reinforcement Learning (RL) policies using simulation models before deployment in real-world environments is a common strategy when r

local-aiarxiv-cs-ai
6 Aug 2026
Model Releases

BIM-Native Tokenization for Constraint-Aware Room Layout Synthesis

DGX agent

arXiv:2512.04832v3 Announce Type: replace Abstract: We present a BIM-native tokenization for room-level layout synthesis in Building Information Modeling (BIM) scenes. The core contribution is represe

model-releasesarxiv-cs-cv
6 Aug 2026
Research

COMPAS: Difficulty-Aware Joint Search for Optimizing Code Generation

DGX agent

arXiv:2608.04336v1 Announce Type: cross Abstract: Code generation systems make each LLM call with a model, a prompt, and decoding settings. However, existing optimization methods usually tune only par

researcharxiv-cs-ai
6 Aug 2026
Research

Elbow-Based MoE Routing: A Training-Free Inference Time Plugin for Expert Selection

DGX agent

arXiv:2608.04401v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models enable model scaling while maintaining low inference-time compute by activating only a subset of experts per token. Howe

researcharxiv-cs-lg
6 Aug 2026
Model Releases

ExeCRE: Execution-Consistency Guided Reliability Estimation for Self-Correcting Code Generation

DGX agent

arXiv:2608.04439v1 Announce Type: cross Abstract: Large language models (LLMs) have made notable progress in code generation, but they still struggle on challenging tasks that require sophisticated al

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

From Score Matrices to Football-Aware Match-State Simulation: An Auditable LLM Harness for Exact-Score Reranking

DGX agent

arXiv:2608.05030v1 Announce Type: new Abstract: Football score forecasting combines a strong statistical core with a difficult contextual edge. Dynamic Poisson-family models estimate team strength, ex

model-releasesarxiv-cs-ai
6 Aug 2026
Research

Fundamentals of quantum Boltzmann machine learning with visible and hidden units

DGX agent

arXiv:2512.19819v2 Announce Type: replace-cross Abstract: One of the primary applications of classical Boltzmann machines is generative modeling, wherein the goal is to tune the parameters of a model

researcharxiv-cs-lg
6 Aug 2026
Model Releases

Hallucinations on the Board: Tool-Augmented Evaluation of LLM Chess Commentary

DGX agent

arXiv:2608.04240v1 Announce Type: cross Abstract: Superhuman game engines in domains like chess have made expert-level evaluations easily accessible, yet they communicate what is true without the natu

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

Long-term Measurements: Towards a Longitudinal Understanding of Human-AI Interactions

DGX agent

arXiv:2608.02491v2 Announce Type: replace Abstract: Language models have taken on the role of a very new type of technology, by virtue of their 'human-ness' and rapid integration into users' daily liv

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

Not Truly Multilingual: Script Consistency as a Missing Dimension in VLM Evaluation

DGX agent

arXiv:2606.17188v3 Announce Type: replace-cross Abstract: Current multilingual evaluations for Vision-Language Models (VLMs) assume a one-to-one mapping between language and orthography, overlooking b

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Predict, Then Retrieve: Cross-Instance Future-State Retrieval from Video Prefixes

DGX agent

arXiv:2608.04426v1 Announce Type: cross Abstract: We introduce Predictive State Retrieval (PSR), a task in which a model observes a short video prefix and a temporal question about an object's future

model-releasesarxiv-cs-cl
6 Aug 2026
← Previous
1…339340341342343…1065
Next →