AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
16,965 results
Model Releases

Tied Trit-Planes: Constraining PTQTP to a Uniform Nine-Level Quantizer, with a Persistent Folded Format for Disk-Streamed Mixture-of-Experts Serving

DGX agent

arXiv:2608.08910v1 Announce Type: new Abstract: PTQTP decomposes LLM weight matrices into two ternary (trit) planes with two free per-group scales. Tying the scales to a fixed ratio of three collapses

model-releasesarxiv-cs-cl
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Time Present and Time Past: Benchmarking Large Language Models on Temporally Evolving Document Understanding

DGX agent

arXiv:2608.08512v1 Announce Type: new Abstract: Evolving documents, such as laws, tax codes, and software documentation, are amended, replaced, and sometimes reverted over time, so a question has diff

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

To Memorize or to Retrieve: Scaling the Interaction Between Pretraining and Retrieval

DGX agent

arXiv:2604.00715v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) improves language model (LM) performance by providing relevant context at test time for knowledge-intensi

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

TomaMMU: A Comprehensive Multimodal Understanding Benchmark for Tomato Leaf Diseases

DGX agent

arXiv:2608.08727v1 Announce Type: cross Abstract: To address this gap, we introduce TomaMMU, a large-scale Tomato leaf disease MultiModal Understanding dataset, alongside TomaBench, a benchmark for ev

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

TongGuOCR: A Layout-Aware and Token-Augmented OCR Framework for Chinese Historical Documents

DGX agent

arXiv:2608.07917v1 Announce Type: new Abstract: Chinese historical documents preserve valuable cultural heritage, but many collections remain accessible only as scanned page images, preventing full-te

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Topology-Aware Global-Local Mamba Networks for Palm Vein Biometrics

DGX agent

arXiv:2608.08951v1 Announce Type: new Abstract: Palm-vein recognition is a fine-grained biometric task in which both local vascular texture and the global layout of the vessel tree carry discriminativ

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Toward Mask Annotation-Free Surgical Instrument Segmentation from Endoscopic Images Using Text-Prompted Segment Anything Model 3 (SAM3)

DGX agent

arXiv:2608.08844v1 Announce Type: new Abstract: Surgical instrument segmentation is a fundamental task for computer-assisted interventions, yet most existing methods rely on pixel-level annotations or

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Toward Metacognitive One-Shot Indirect Prompt Injection: Strategy Abstraction Via Outcome-Conditioned Reflection

DGX agent

arXiv:2608.08795v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents are vulnerable to indirect prompt injection (IPI), in which malicious instructions embedded in external o

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Towards an LLM-based method for quantifying the sexual content in song lyrics

DGX agent

arXiv:2608.08885v1 Announce Type: cross Abstract: Reggaeton is one of the most widely consumed music genres in the world, and its lyrics are commonly regarded as highly sexualized. This claim rests mo

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Towards Expert-level Medical AI for Real-time Video Consultations

DGX agent

arXiv:2608.09861v1 Announce Type: new Abstract: Audio-visual interaction is the standard for patient-physician consultations, enabling natural communication and effective assessment of illness through

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Towards Researcher Agents for Knowledge-Graph Question Answering

DGX agent

arXiv:2608.07700v1 Announce Type: new Abstract: Translating a natural-language question into a SPARQL query that can be executed against a large knowledge graph requires resolving lexical ambiguity, g

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

TRACE: TRajectory Attribution for Automated Context Engineering

DGX agent

arXiv:2608.09153v1 Announce Type: new Abstract: Production AI agents fail when their context sources -- system prompts, knowledge bases, tool descriptions, and procedural skills -- contain errors or g

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Tracking the Best Strategy in an Extensive-Form Game

DGX agent

arXiv:2608.09501v1 Announce Type: new Abstract: We consider the extensive-form bandit problem where on each trial the learner plays an extensive-form game against an oblivious adversary. We focus on t

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Trajectory Design and Budgeted Querying for Digital Twin Calibration

DGX agent

arXiv:2608.08631v1 Announce Type: new Abstract: Digital-twin calibration requires interaction data that is expensive to collect. We study two acquisition decisions: which trajectories to generate, and

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Trajectory Divergence Horizon Decision for Reliable Dual-Arm Surgical Subtask Manipulation

DGX agent

arXiv:2608.09125v1 Announce Type: new Abstract: Surgical robotic systems are increasingly being adopted as clinical workload rises, motivating autonomous solutions for repetitive manipulation subtasks

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

TREAT: Evaluating Access to Formal Knowledge across Equivalent Mathematical Representations

DGX agent

arXiv:2608.07540v1 Announce Type: new Abstract: AI systems increasingly operate between flexible input representations and formal objects used by downstream tools. A key challenge is recognizing when

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

TreeHop: Efficient Embedding-Level Query Rewriter

DGX agent

arXiv:2504.20114v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems face significant challenges in multi-hop question answering (MHQA), where complex queries require

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

TrustRoboReward: Preference-Ordered Isotonic Score Editing for Multi-Paradigm Robot Reward Models

DGX agent

arXiv:2608.08491v1 Announce Type: new Abstract: Reward models are a bottleneck for reinforcement learning in embodied AI. Long-horizon robotic manipulation requires scalable vision feedback beyond han

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Two-Layer Linear Auto-Regressive Models Estimate Latent States

DGX agent

arXiv:2606.12691v2 Announce Type: replace-cross Abstract: Auto-regressive models have emerged as powerful tools for sequential data, from language to video. Understanding how and why these models lear

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Understanding Calibration and Truncation Error Propagation in Training-Free Low-Rank Compression for LLMs

DGX agent

arXiv:2608.08506v1 Announce Type: new Abstract: Training-free low-rank compression frameworks have been gaining prominence for LLM compression given their effectiveness in reducing model parameter cou

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Unified Hallucination Fuzzing for Multimodal Large Language Models

DGX agent

arXiv:2608.07525v1 Announce Type: cross Abstract: Hallucination remains a persistent challenge for Multimodal Large Language Models (MLLMs), severely limiting their reliability in high-stakes applicat

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

UniMoMo: Expert Merging-Based MoE Acceleration for Large Recommendation Models

DGX agent

arXiv:2608.08627v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) layers expand recommendation capacity through conditional computation, yet a trained checkpoint still stores and routes

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

UNMASK: Discovering and Causally Verifying Spurious Shortcuts in Text Classifiers

DGX agent

arXiv:2608.09209v1 Announce Type: new Abstract: Neural language models trained on large crowdsourced corpora frequently exploit spurious surface patterns tied to target labels without true linguistic

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

UNSPECIFIC: General Constraint Synthesis for Breaking Copy-and-Paste Shortcut in LLM Instruction Following

DGX agent

arXiv:2608.09154v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly expected to follow long lists of constraints in complex instructions, and synthesizing instructions from a

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Unsupervised Point Cloud Registration with Self-Distillation

DGX agent

arXiv:2409.07558v2 Announce Type: replace Abstract: Rigid point cloud registration is a fundamental problem and highly relevant in robotics and autonomous driving. Nowadays deep learning methods can b

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

VCU-Bridge: Hierarchical Visual Connotation Understanding via Semantic Bridging

DGX agent

arXiv:2511.18121v2 Announce Type: replace-cross Abstract: While Multimodal Large Language Models (MLLMs) excel on benchmarks, their processing paradigm differs from the human ability to integrate visu

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool Use

DGX agent

arXiv:2608.08477v1 Announce Type: new Abstract: We present VectraYX-Vision-1B, a sub-2B vision-language model (VLM) for Spanish/LATAM cybersecurity imagery, coupling a frozen SigLIP-so400m encoder to

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

VeinCast: Physics-Guided Dynamic Field Graphs with Graph-Conditioned Fusion for Global Medium-Range Weather Forecasting

DGX agent

arXiv:2608.09286v1 Announce Type: cross Abstract: Global medium-range weather forecasting requires modeling structured yet state-dependent interactions among heterogeneous atmospheric fields. Existing

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Verication-driven closed-loop multi-agent large language modelframework for code-compliant structural design

DGX agent

arXiv:2608.07978v1 Announce Type: cross Abstract: Multi-agent large language model(LLM)systems are applied to structural design,yet most use one-shot generation and cannot verify their output,leaving

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

VideoVIBE: A Video-Grounded Diagnostic Benchmark for One-Shot Interactive Website Generation

DGX agent

arXiv:2608.09573v1 Announce Type: new Abstract: Natural-language-driven 'vibe coding' enables the one-shot generation of visually rich and interactive web applications, yet reliable assessment of thei

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

VIGIL: Tackling Hallucination Detection in Image Recontextualization

DGX agent

arXiv:2602.14633v2 Announce Type: replace Abstract: We introduce VIGIL (Visual Inconsistency & Generative In-context Lucidity), a benchmark dataset and framework that provides a fine-grained categoriz

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

VLZip: Unified Visual and Textual Compression for Interleaved Long-Context Modeling

DGX agent

arXiv:2608.08630v1 Announce Type: new Abstract: Vision Language Models (VLMs) face significant challenges with ultra-long, interleaved image-text sequences due to the quadratic complexity of self-atte

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

VTO: Visual Tool Orchestration for Video Anomaly Detection

DGX agent

arXiv:2608.08219v1 Announce Type: cross Abstract: Video anomaly detection (VAD) is a critical yet challenging task due to the complex and diverse nature of real-world scenarios. Traditional deep learn

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Weak Correlations as the Underlying Principle for Linearization of Gradient-Based Learning Systems

DGX agent

arXiv:2401.04013v2 Announce Type: replace Abstract: Deep learning models, such as wide neural networks, can be conceptualized as nonlinear dynamical physical systems characterized by a multitude of in

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

WebChoreArena: Evaluating Web Browsing Agents on Realistic Tedious Web Tasks

DGX agent

arXiv:2506.01952v2 Announce Type: replace-cross Abstract: Powered by large language models (LLMs), web browsing agents operate graphical user interfaces in a human-like manner, offering a transparent

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

What to Edit Next: Visually Aligned Image-Editing Follow-Up Suggestions in Conversational Systems

DGX agent

arXiv:2608.07565v1 Announce Type: cross Abstract: Conversational assistants increasingly recommend follow-up edits to help users continue a task. Existing systems primarily target text-only interactio

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

What Would Fix This RAG Failure? Auditing Counterfactual Response with Paired Evidence Interventions

DGX agent

arXiv:2608.08944v1 Announce Type: cross Abstract: A failed retrieval-augmented generation (RAG) answer can be consistent with several unseen responses to evidence repair. We introduce Pair-ID, an offl

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Counterbalancing Hides the Bias: Access-Conditioned Position Lock in Forced-Choice LLM Evaluation

DGX agent

arXiv:2607.10202v2 Announce Type: replace Abstract: Forced-choice probes with counterbalanced orientations are a standard tool for measuring language-model 'value dispositions,' and a concentration/ex

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Do Task Vectors Interfere? Mapping the Validity Boundaries of Weight-Space Composition

DGX agent

arXiv:2608.09490v1 Announce Type: new Abstract: Task arithmetic treats fine-tuning displacements as composable directions in weight space, yet it remains unclear when parameter addition reflects predi

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Does An Extra View Help? Adapting Single-View 3D Reconstruction with Extra Imagery

DGX agent

arXiv:2608.08132v1 Announce Type: new Abstract: Reconstruction of 3D objects from a single image is a challenging research problem in computer vision. The key challenge is the lack of critical informa

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output

DGX agent

arXiv:2503.24191v4 Announce Type: replace-cross Abstract: Content Warning: This paper may contain unsafe or harmful content generated by LLMs that may be offensive to readers. Large Language Models (L

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When Is Benchmark Contamination Detectable? Information Limits and Power-Calibrated Audits

DGX agent

arXiv:2608.07914v1 Announce Type: new Abstract: Behavioral contamination detectors can return 'no evidence' either because a benchmark is clean or because the audit has little power. We formalize this

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When LLM Agents Negotiate: Private Information and Dynamic Bargaining in Supply Chains

DGX agent

arXiv:2608.07538v1 Announce Type: new Abstract: As LLM agents move from decision support to autonomous procurement, firms need to know whether delegated negotiators create value, divide it predictably

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When should we trust the annotation? Selective prediction for molecular structure retrieval from mass spectra

DGX agent

arXiv:2603.10950v2 Announce Type: replace Abstract: Machine learning methods for identifying molecular structures from tandem mass spectra (MS/MS) have advanced rapidly, yet current approaches still e

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Skills Meet Safety: Benchmarking and Characterizing the Adaptive Jailbreak Robustness of Skill-Merged LLMs

DGX agent

arXiv:2608.08542v1 Announce Type: new Abstract: Model merging has become the default way to give an aligned language model new skills without retraining: a practitioner folds task vectors from math, c

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines

DGX agent

arXiv:2608.07813v1 Announce Type: new Abstract: An LLM judge deployed inside a reasoning pipeline does not merely measure quality, it decides which answer ships. We show that the cost of that decision

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation

DGX agent

arXiv:2608.07762v1 Announce Type: new Abstract: LLM benchmarks can build an organization's reputation and attract customers, but only when results are transparent and verifiable. Unverified claims tha

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Models

DGX agent

arXiv:2608.00591v2 Announce Type: replace Abstract: A calibrated stochastic world model can reveal how uncertain a future is without revealing why it branches. The same conditional future law can aris

model-releasesarxiv-cs-ai
11 Aug 2026
← Previous
1…1011121314…354
Next →