AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
All
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,055 results
Model Releases

SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information

DGX agent

arXiv:2608.10692v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as mobile assistants, where a key challenge is leveraging personal information scattered across

model-releasesarxiv-cs-ai
12 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Static in Frames, Dynamic in Events: Rethinking Features in Event Cameras as Motion Cues

DGX agent

arXiv:2608.11075v1 Announce Type: new Abstract: Event cameras capture intensity changes asynchronously with high temporal resolution, requiring novel preprocessing methods for downstream tasks. Unlike

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generation

DGX agent

arXiv:2608.10439v1 Announce Type: new Abstract: Streaming video generation holds strong potential for world modeling, where future frames must be inferred online sequentially to form a continuous vide

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

Surfacing the Unsaid: CUE-Bench for Affective Stance in Chinese Discourse

DGX agent

arXiv:2608.10810v1 Announce Type: cross Abstract: Emotion understanding in discourse requires reasoning beyond surface sentiment because speakers often convey affect through indirect, implicit, polite

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

TACTICL: Task-Aware Compression of Tabular ICL Models

DGX agent

arXiv:2608.10837v1 Announce Type: cross Abstract: The strong performance of foundation models for tabular tasks comes at substantial inference costs. Distilling models into task-specific architectures

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

TAF-MED: Multi-Turn Safety Refusal Collapse in LLMs Under Declared Self-Treatment Intent

DGX agent

arXiv:2608.10258v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly provide conversational health information that may influence treatment decisions, yet existing benchmarks do

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

TEASR: Training-Efficient Any-Step Diffusion Transformer for Real-World Image Super-Resolution

DGX agent

arXiv:2606.16188v2 Announce Type: replace Abstract: Diffusion models excel in Real-World Image Super-Resolution (Real-ISR) due to their powerful generative priors but suffer from slow iterative sampli

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

TemMed-Bench: Evaluating Temporal Medical Image Reasoning in Vision-Language Models

DGX agent

arXiv:2509.25143v2 Announce Type: replace-cross Abstract: Existing medical reasoning benchmarks for vision-language models primarily focus on analyzing a patient's condition based on an image from a s

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

Temporally Grounded Compositional Camera Motion Understanding via Geometric Knowledge Distillation

DGX agent

arXiv:2608.10932v1 Announce Type: cross Abstract: Understanding camera motion is fundamental to video perception, with applications in spatial intelligence and controllable video generation. Multimoda

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Test-Time Self-Evolving GUI Visual Grounding via Reflection-Guided On-Policy Self-Distillation

DGX agent

arXiv:2608.11191v1 Announce Type: cross Abstract: GUI Visual Grounding is a fundamental capability for GUI agents. Existing models typically freeze their parameters after deployment, limiting their ab

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Tested Nemotron 3.5 Lightning locally on coding, Hermes Agent and agentic work

DGX agent

Ran the model with quants (Q5) and MTP by bartowski with llama.cpp server. It takes ~24GB ram running on M5 Pro with 48GB at about 65t/s. On some tasks it was quite the overthinker. Overall, the quali

model-releasesr-localllama
12 Aug 2026
Model Releases

The Evaluation Protocol Determines the Result: An Independent Reproduction of LeWorldModel on TwoRoom

DGX agent

arXiv:2608.10145v1 Announce Type: new Abstract: LeWorldModel trains a latent world model with a prediction loss and a single anti-collapse regulariser, and reports approximately 87% of goals reached o

model-releasesarxiv-cs-lg
12 Aug 2026
Model Releases

The Gaussian-Multinoulli Restricted Boltzmann Machine: A Potts Model Extension of the GRBM

DGX agent

arXiv:2505.11635v2 Announce Type: cross Abstract: Many real-world tasks, from associative memory to symbolic reasoning, benefit from discrete, structured representations that standard continuous laten

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

The Multilingual Quantization Tax: Structural Collapse and Typological Fragility in Edge SLMs

DGX agent

arXiv:2608.09941v1 Announce Type: new Abstract: While 4-bit weight quantization is critical for deploying Small Language Models (SLMs) on edge devices, evaluations of the resulting performance degrada

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

The Truth Stays in the Family: Enhancing Contextual Grounding via Inherited Truthful Heads in Model Lineages

DGX agent

arXiv:2606.15821v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have produced many specialized multimodal LLMs (MLLMs) that share common foundational LLMs, fo

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Toward Human Rights Benchmarking for LLMs: A Pilot Methodology

DGX agent

arXiv:2608.10268v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate legal determinations over what human rights are realized, and how. Yet, no evaluation benchmark exis

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Towards Efficient Reasoning in LLM-Based Recommender Systems via Model Merging

DGX agent

arXiv:2608.10447v1 Announce Type: cross Abstract: Large language model-based recommender systems are increasingly adopting slow-thinking models that generate step-by-step reasoning before making predi

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Towards Unified Dynamic Face Landmark Detection

DGX agent

arXiv:2608.10346v1 Announce Type: cross Abstract: Although advancements in face landmark detection (FLD) methods continue to push performance boundaries, they overlook two major functional limitations

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

TRACE: Trustworthy Retrieval-Augmented Conversational Engine

DGX agent

arXiv:2608.10176v1 Announce Type: new Abstract: Public service chatbots are expected to deliver recommendations from an underlying public service directory, while also making sure that the recommendat

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Try Qwen-Image-3.0 on @openart_ai! 🎨👀

DGX agent

Try Qwen-Image-3.0 on @openart_ai! 🎨👀 Qwen Image 3.0 is now on OpenArt ✨ The most Real Qwen image model yet. Native text across 12 languages, precise 10px type, and full interfaces like web pages, gam

model-releasesqwen--x
12 Aug 2026
Model Releases

Uncertainty-Aware Ensemble Deep Randomized Neural Networks for Classification

DGX agent

arXiv:2608.10007v1 Announce Type: cross Abstract: The current state-of-the-art (SOTA) deep randomized neural networks, such as deep Random Vector Functional Link (dRVFL) and ensemble deep RVFL (edRVFL

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs

DGX agent

arXiv:2608.10042v1 Announce Type: cross Abstract: Tool-use LLMs are increasingly asked to act on users' behalf, but existing benchmarks usually focus on profile recall, style imitation, generic tool u

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference

DGX agent

arXiv:2603.18446v2 Announce Type: replace Abstract: Long-context inference remains challenging for large language models due to attention dilution and out-of-distribution degradation. Context selectio

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

V-FiLLM: Verified Financial LLM Reasoning Benchmark

DGX agent

arXiv:2608.11047v1 Announce Type: new Abstract: While existing benchmarks have made substantial progress in evaluating LLMs across STEM domains, financial reasoning over structured data remains compar

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?

DGX agent

arXiv:2608.10875v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as personal assistants. Existing evaluations, however, mostly use short, self-contained re

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

VisEditBench: Can Vision-Language Models Edit Visualization Code from Multimodal Feedback?

DGX agent

arXiv:2608.10408v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong capabilities in generating visualization code from textual or visual specifications. However, real-world

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

Vision-Language-Motion Maps: An Open-Vocabulary, Uncertainty-Aware, Queryable Motion Attribute for 3D Scene Maps

DGX agent

arXiv:2607.16173v2 Announce Type: replace Abstract: Open-vocabulary 3D maps let robots answer language queries about what and where, but they assume a static world and cannot answer queries about how

model-releasesarxiv-cs-ro
12 Aug 2026
Model Releases

Visual Geometry Foundation-Aware Gaussians for Single-Frame Surround-View Driving Reconstruction

DGX agent

arXiv:2608.10682v1 Announce Type: new Abstract: Single-frame surround-view reconstruction faces severe geometric instability and rendering artifacts due to minimal inter-camera overlap. While existing

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

VoxSumm: A Multilingual Corpus of Long-Form Spoken News for Joint Summarization and Translation

DGX agent

arXiv:2608.10359v1 Announce Type: cross Abstract: As information increasingly traverses linguistic boundaries, users require concise cross-lingual representations of long-form content. Nevertheless, l

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

We have her dash camera which shows they are lying. The agents are wearing body cameras and should have dash cameras of their own. If what t…

DGX agent

We have her dash camera which shows they are lying. The agents are wearing body cameras and should have dash cameras of their own. If what they say happened was true they wouldn’t be issuing statement

model-releasesanthropic--x
12 Aug 2026
Model Releases

What unique, custom QOL upgrades have you given your local agents?

DGX agent

Warning: Kinda long post. If you don't like reading, please skip for your own sanity. Also, I've got nothing to sell, just a tinkerer, so I just want to share ideas and learn from you guys too. When I

model-releasesr-localllama
12 Aug 2026
Model Releases

When Chain-of-Thought Helps and When It Hurts: An Empirical Investigation of the Serial-Depth Bottleneck in LLM Reasoning

DGX agent

arXiv:2608.09942v1 Announce Type: cross Abstract: It is widely assumed that chain-of-thought (CoT) prompting universally improves LLM reasoning. We investigate this through the conceptual framework of

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models

DGX agent

arXiv:2608.11024v1 Announce Type: new Abstract: Attribute hallucination---where vision-language models (VLMs) correctly identify an object but mischaracterize its properties---is prevalent yet mechani

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

Whole-Body Planning for Humanoids Navigating Confined Spaces via Self-Collision Avoidance References

DGX agent

arXiv:2608.10220v1 Announce Type: new Abstract: Humanoid locomotion in highly confined environments requires navigating dense environmental obstacles and complex self-collision bounds while maintainin

model-releasesarxiv-cs-ro
12 Aug 2026
Model Releases

Why Does CLAUDE.md Keep Growing? Catastrophic Remembering in Agentic Coding

DGX agent

arXiv:2608.11095v1 Announce Type: new Abstract: Agentic coding READMEs like CLAUDE.md grow without bound in real repositories, stopping only when the repository retires or someone rewrites the file wh

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Withholding the Completing Chunk: Deterministic Pair-Completion Guardrails for Streaming LLM Output

DGX agent

arXiv:2608.10279v1 Announce Type: cross Abstract: Streaming language-model output creates a release-timing problem: complete-response moderation acts after streamed text has escaped, whereas repeated

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Workflow Cards: Structured Summaries of Workflow Executions Using Provenance Data

DGX agent

arXiv:2608.11022v1 Announce Type: cross Abstract: Model Cards and Data Cards have demonstrated the value of structured, human-readable documentation for machine learning artifacts, capturing their con

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

1 Day in and I feel okay saying Muse-Glimmer-30B finally beats 3.6-27B for the size in some use-cases

DGX agent

A few things right off the bat: it reasons very efficiently. Like Grok 4.5 levels of efficient thinking it quantizes very well. My first few tests with iq3_xxs were better than Qwen/Gemma behaved at t

model-releasesr-localllama
11 Aug 2026
Model Releases

10 year garbage card for local llms

DGX agent

Hello everyone! ​I like dumb things. I like working with weak computers and microcontrollers. I like the simplicity and low electricity usage. Simply put, the efficiency of a 'dumb' PC. ​The first tim

model-releasesr-localllama
11 Aug 2026
Model Releases

12GB VRAM gang, what's our plan?

DGX agent

Seems like we're limited to qwen finetuned MoEs for now. Looking at the current landscape - focus seems to be on dense models (muse glimmer 30b, qwen 3.8 27b) for smaller setups. Is upgrading to 24GB

model-releasesr-localllama
11 Aug 2026
Model Releases

360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents

DGX agent

arXiv:2608.08814v1 Announce Type: cross Abstract: We present 360CityArena, a benchmark for evaluating the urban exploration capabilities of embodied agents within a photorealistic environment construc

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

⚡️A coalition that secures long-term AI capacity: We’re aggregating long-term compute demand in Europe to determine what capacity is built, …

DGX agent

⚡️A coalition that secures long-term AI capacity: We’re aggregating long-term compute demand in Europe to determine what capacity is built, where it’s located, and whom it serves. Through these multi-

model-releasesmistral-ai--x
11 Aug 2026
Model Releases

A Control Function Framework for Mitigating Position Bias in Learning to Rank Systems

DGX agent

arXiv:2506.06989v3 Announce Type: replace-cross Abstract: Learning-to-rank (LTR) systems commonly depend on implicit feedback, such as user clicks, because it is easy to collect and can serve as a val

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

A Fair Objective for Human-Empowerment-Preserving AI: Desiderata, Design, and Likely Behavioral Consequences

DGX agent

arXiv:2608.08240v1 Announce Type: new Abstract: This paper explores the idea of promoting well-being and safety in human-AI interactions by forcing AI agents explicitly to empower humans and to manage

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

A Rigorous Turing Test: a Foundation for Evaluating Artificial General Intelligence

DGX agent

arXiv:2501.17629v2 Announce Type: replace-cross Abstract: Several studies claim that large language models have passed the Turing Test and hence can 'think', yet none follow Turing's original instruct

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

A Tight Lower Bound for Smooth Nonconvex Stochastic Optimization with Bounded Gradient Noise

DGX agent

arXiv:2608.09004v1 Announce Type: cross Abstract: We prove a sharp lower bound for smooth nonconvex stochastic optimization with uniformly bounded gradient noise. In the (K=1) fresh-sample model, ever

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

A Unified Issue Resolution Benchmark for Requirement Clarification, Planning, and Code Generation for Coding Agents

DGX agent

arXiv:2608.09072v1 Announce Type: cross Abstract: Large language model-powered coding agents are increasingly used to modify existing code repositories, for example, by adding features or fixing bugs.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Accelerate PostgreSQL migrations using Gemini in Database Migration Service

DGX agent

Imagine this scenario: Your team decides to migrate a core application from an existing commercial database like Oracle or SQL Server to open source PostgreSQL or a fully managed service such as Alloy

model-releasesgoogle-cloud-ai
11 Aug 2026
← Previous
123456…460
Next →