AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
All
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,106 results
Model Releases

ICM Out! Better Tournament Strategy from Computed Continuations, vs. Solvers and LLMs

DGX agent

arXiv:2608.09586v1 Announce Type: new Abstract: The Independent Chip Model (ICM) converts tournament chips into reference prize equity, and policies are routinely constructed against those values. Bec

model-releasesarxiv-cs-ai
11 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Illusion or Integrity? Geometrical Consistency Metric for AIGC Video Quality Evaluation

DGX agent

arXiv:2608.09594v1 Announce Type: cross Abstract: Recently, AI-driven video generation has attracted considerable attention. This surge increases the demand for reliable video quality assessment (VQA)

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Implicit Virtual Leader: Decentralized Vision-Only Relative Pose Estimation for Multi-Robot Formations

DGX agent

arXiv:2607.15708v2 Announce Type: replace Abstract: Classical leader-follower formation control suffers from single points of failure and error propagation, and relies on absolute localization sensors

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

IndexTTS 2.5 Technical Report

DGX agent

arXiv:2601.03888v4 Announce Type: replace-cross Abstract: In prior work, we introduced IndexTTS 2, a zero-shot neural text-to-speech foundation model comprising two core components: a transformer-base

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

InfoOps Bench: A live information operations safety benchmark

DGX agent

arXiv:2607.28503v3 Announce Type: replace Abstract: In this paper we present an active, constantly updated AI benchmark which measures the integrity of frontier language models against being co-opted

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Instability of LLM Pre-Pretraining: It Doesn't Always Help. An Investigation on Multiple Languages

DGX agent

arXiv:2608.08800v1 Announce Type: new Abstract: Pretraining LLMs on artificial languages ('pre-pretraining') is a technique that could reportedly increase token efficiency by 33%, i.e., save up to 33%

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Instruction Set and Language for Hypergraphs

DGX agent

arXiv:2607.10194v2 Announce Type: replace-cross Abstract: We present IsalHG, a method for representing the structure of any finite, connected hypergraph of bounded hyperedge arity as a string over a c

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

InstructionCrafter: Generating Consistent and High-Fidelity Visual Instructions

DGX agent

arXiv:2608.08460v1 Announce Type: new Abstract: Given textual task instructions, generating step-by-step visual instructions as an image sequence requires the simultaneous satisfaction of multiple pro

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models a…

DGX agent

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models are pushing the frontier of coding and knowledge work, but su

model-releasesjerry-liu--x
11 Aug 2026
Model Releases

Introducing 𝗘𝘅𝘁𝗿𝗮𝗰𝘁𝗕𝗲𝗻𝗰𝗵: the most comprehensive benchmark for information extraction from complex enterprise documents. Our app…

DGX agent

Introducing 𝗘𝘅𝘁𝗿𝗮𝗰𝘁𝗕𝗲𝗻𝗰𝗵: the most comprehensive benchmark for information extraction from complex enterprise documents. Our applied research team tested: 14 systems — frontier VLMs, coding agents, ex

model-releasesjerry-liu--x
11 Aug 2026
Model Releases

Introducing Unsloth Desktop app

DGX agent

Hi LocalLlama, we're super excited to release Unsloth Desktop today! 🦥 It's the first desktop app that enables you to run and train models locally. Open-source. Available on Mac, Windows, and Linux Su

model-releasesr-localllama
11 Aug 2026
Model Releases

Is the ACL Responsible NLP Checklist a Box-Ticking Exercise? A Large-Scale Analysis of EMNLP 2025

DGX agent

arXiv:2608.09280v1 Announce Type: new Abstract: Responsible NLP practice includes a) transparency, b) ethics, and c) societal impacts. The Responsible NLP Checklist aims to push these goals, and promo

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

It's been exciting for Ollama to partner with @JensenHuang and the @NVIDIAAI team on launching open models. Open models have no boundaries, …

DGX agent

It's been exciting for Ollama to partner with @JensenHuang and the @NVIDIAAI team on launching open models. Open models have no boundaries, and let's continue to work together to make this ecosystem b

model-releasesollama--x
11 Aug 2026
Model Releases

Jako Tako or Fluent? Presenting PoVisLE: A Polish Vision-Language Evaluation

DGX agent

arXiv:2608.07763v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong performance on tasks such as image captioning, visual question answering, and image-to-text generatio

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

JUMP-lite: Compact, reproducible benchmarking of cell representations

DGX agent

arXiv:2608.07632v1 Announce Type: cross Abstract: Image-based profiling captures rich phenotypic signatures for drug discovery and functional genomics. Large public datasets like JUMP Cell Painting no

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Kernel Methods for Refined Prophet Inequalities

DGX agent

arXiv:2608.08662v1 Announce Type: cross Abstract: The single-selection prophet inequality is a canonical Bayesian online selection problem in which independent nonnegative values arrive sequentially a

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

KGCaRe: Explainable Complex Conditional Question Answering using Automatic Knowledge Graph Construction and Context Retrieval with LLMs

DGX agent

arXiv:2608.09779v1 Announce Type: cross Abstract: Answering complex conditional questions using Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) remains a challenge, particularly

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Knowing You Is Everything: LLM Agents Achieve Near-Perfect Profile-Consistent Reaction Prediction in Social Media Simulation

DGX agent

arXiv:2608.07498v1 Announce Type: cross Abstract: Autonomous AI agents in social media present concrete risks to democratic discourse and platform governance, while also offering tools for pre-deploym

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

KVDiagnosis: A Diagnostic Benchmark for KV-Cache Compression in Long-Context Language Models

DGX agent

arXiv:2608.09412v1 Announce Type: new Abstract: KV-cache compression reduces long-context memory, but aggregate task scores reveal neither which correct executions fail nor why. We present KVDiagnosis

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Label-Free Parkinson's Disease Screening from Face and Voice through Mechanistic Interpretability

DGX agent

arXiv:2608.08976v1 Announce Type: new Abstract: Parkinson's disease (PD) is the second most common neurodegenerative disorder. Typical machine learning screening methods require PD labels, but the ava

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Large Language Models Align with the Human Brain during Creative Thinking

DGX agent

arXiv:2604.03480v2 Announce Type: replace-cross Abstract: Creative thinking is a fundamental aspect of human cognition, and divergent thinking-the capacity to generate novel and varied ideas-is widely

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Large Multimodal Agents for Intelligent Transportation Systems: Architectures, Evidence, and Deployment Challenges

DGX agent

arXiv:2608.08184v1 Announce Type: new Abstract: Large multimodal agents (LMAs) are increasingly proposed for intelligent transportation systems (ITS), but existing studies often conflate multimodality

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LazyHMC: Hamiltonian Monte Carlo Simulation for Lazy, Infinite Dimensional Probabilistic Programs

DGX agent

arXiv:2608.08588v1 Announce Type: cross Abstract: Hamiltonian Monte Carlo (HMC) is a successful generic inference method in probabilistic programming, but in its ordinary formulation it needs gradient

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Learning How the World Evolves: Extrapolative Video World Models via Latent Dynamics Reasoning

DGX agent

arXiv:2608.09926v1 Announce Type: new Abstract: The world evolves following its dynamics, i.e., its laws of motion. However, leading video diffusion models largely fit the pixels without modeling how

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Learning human joint torques from pixels

DGX agent

arXiv:2608.09083v1 Announce Type: new Abstract: Estimating human joint torques from visual observations is a key step toward bringing biomechanical analysis from controlled laboratories to real-world

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Learning Structural Illumination for Unsupervised Low-light Enhancement

DGX agent

arXiv:2608.08153v1 Announce Type: new Abstract: Existing unsupervised low-light image enhancement (LLIE) methods often estimate illumination directly from the entire low-light input, without separatin

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Learning to Triage Vulnerability Reports from Program Analysis: An Empirical Study in Node.js

DGX agent

arXiv:2510.20739v2 Announce Type: replace-cross Abstract: Program analysis tools often produce large volumes of candidate vulnerability reports that require costly manual review, creating a practical

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?

DGX agent

arXiv:2503.19990v4 Announce Type: replace Abstract: Many real-world applications of spatial intelligence, such as robotic control, autonomous driving, and automated assembly, require spatial reasoning

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LegoLM: Structured Weight Sharing for Large Language Models

DGX agent

arXiv:2608.08652v1 Announce Type: cross Abstract: We present LegoLM{}, a structured weight-sharing compression framework for large language models grounded in a systematic study of why global weight s

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LexKairos: Benchmarking Legal Temporal Capabilities in LLMs

DGX agent

arXiv:2608.09106v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated strong performance across a wide range of legal tasks. In legal practice, time is a critical concept that

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

LGNNIC: Acceleration of Large-Scale GNN Training using SmartNICs

DGX agent

arXiv:2608.07733v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are widely used across domains such as natural sciences, social network analysis, chip design, and recommendation systems

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LIBAD: A Multimodal Anomaly Detection Benchmark for Li-Ion Battery Electrode Manufacturing

DGX agent

arXiv:2608.07958v1 Announce Type: new Abstract: Multimodal industrial anomaly detection has largely focused on discrete products using strongly correlated RGB and 3D observations, leaving continuous p

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

LIRA: Local Cross-Layer Information Routing for Vision-Language-Action Decoding

DGX agent

arXiv:2608.07596v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models transform representations from pretrained vision-language models (VLMs) into robot actions, yet the interface that

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Listen, See and Track: Spatio-Temporal Audio-Visual Sound Event Reasoning for Omni-Modal Language Models

DGX agent

arXiv:2608.09435v1 Announce Type: new Abstract: Understanding dynamic sound sources requires jointly determining what produces a sound, where the source is located, and how it moves over time. Yet exi

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Listwise Cross-Encoder Fine-Tuning vs. Agentic Instruction Tuning for LLM Rerankers: A Systematic Study in Medical Procedure Reranking

DGX agent

arXiv:2608.09650v1 Announce Type: cross Abstract: Reranking medical procedures against patient queries is a critical component of health insurance information retrieval, complicated by a substantial l

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Llama-CPP Parallel Agents --> fine for decode, but one agent's prefill will grind all other agents to a halt

DGX agent

Testing with 3-5 agents. Decode performance is superb, however if one performs a web search and needs to process a few thousand tokens, ALL other agents will grind to a halt: I've tried tuning a littl

model-releasesr-localllama
11 Aug 2026
Model Releases

[llama.cpp PR #26608] Ling-3.0 support (unmerged)

DGX agent

aetherbird has done some great work getting Ling-3.0 to work in llama.cpp. The architecture is generally identical to deepseekv2. I recently added a microscopic 40 line PR to his that adds support for

model-releasesr-localllama
11 Aug 2026
Model Releases

LLM-Driven AutoML for Cross-Lingual Handwritten OCR: Closed-Loop Neural Architecture Search with GPT-5, GPT-4o, and Claude Sonnet 4

DGX agent

arXiv:2607.15509v2 Announce Type: replace-cross Abstract: We present a fully automated closed-loop AutoML framework that uses GPT-5, GPT-4o, and Claude Sonnet 4 as autonomous neural architecture desig

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LLM-Guided Heuristic Design from Simulation Traces: A Case Study in Dynamic Production and AGV Scheduling

DGX agent

arXiv:2608.09343v1 Announce Type: new Abstract: Simulation-based optimization (SBO) evaluates executable policies under stochastic dynamics, but most methods treat the simulator as a black box: aggreg

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LLM within MCP Matters: Measuring Inefficient Resource Utilization Driven by LLMs

DGX agent

arXiv:2608.08467v1 Announce Type: new Abstract: The Model Context Protocol (MCP) standardizes how servers expose data and tools to Large Language Models (LLMs). A common server design embeds frequentl

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LLMs still produce bugs, but those bugs are different than what they used to be. It’s less off-by-ones and more about system design, ui usab…

DGX agent

LLMs still produce bugs, but those bugs are different than what they used to be. It’s less off-by-ones and more about system design, ui usability, missing broader context. Some kinds of coding has bee

model-releasesboris-cherny--x
11 Aug 2026
Model Releases

LLMVisor: A Real-Time Latency Attribution Model for Multi-Tenant LLM Serving

DGX agent

arXiv:2608.08382v1 Announce Type: new Abstract: As LLM inference shifts to multi-tenant GPU clusters, co-batching improves throughput but obscures per-tenant usage and limits control. Enabling fractio

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Local Benchmark : Muse Glimmer 30B vs Qwen 3.6 27B vs Gemma4 31B (and many other models and finetunes)

DGX agent

Needs a lot of requests compared to Qwen (almost twice) and Gemma (almost x3). Final score is fine, even though it is 'not a coding model' https://wonderrico.github.io/local_llm_benchmark/benchmark-ma

model-releasesr-localllama
11 Aug 2026
Model Releases

Locating Failure in Multi-Page Visually Rich Document Understanding: An Empirical Attribution

DGX agent

arXiv:2608.07943v1 Announce Type: new Abstract: Multi-page visually-rich document understanding (MP-VRDU) requires managing evidence that is sparse, spread across pages, and often exceeds a model's co

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LogicIF: Towards Complex Logic Instruction Following

DGX agent

arXiv:2508.09125v4 Announce Type: replace Abstract: Instruction following has catalyzed the recent era of Large Language Models (LLMs) and is the foundational skill underpinning more advanced capabili

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

LogiShot: Logically Coherent Cross-Shot Video Generation

DGX agent

arXiv:2608.08820v1 Announce Type: new Abstract: Generating cross-shot videos that are logically connected is essential for content creation. Currently, most cross-shot video-generation workflows, such

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Long SKILL Compliance as Logical Reasoning: Closure-Grounded Detection with Scaling-Guided On-Policy Distillation

DGX agent

arXiv:2608.08146v1 Announce Type: new Abstract: The increasing complexity of enterprise business scenarios has promoted the widespread adoption of long SKILL documents in agent systems, posing new cha

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Looker’s semantic layer governs Gemini Enterprise data for user trust

DGX agent

For organizations deploying AI agents at scale, there’s often a critical divide between structured and unstructured data. While large language models (LLMs) excel at parsing text documents, emails, an

model-releasesgoogle-cloud-ai
11 Aug 2026
← Previous
1…910111213…461
Next →