AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlog
88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,694 results
Model Releases

Benchmarking and Enhancing LLMs for Rule-Intensive Review of National Standard Documents

DGX agent

arXiv:2608.06312v1 Announce Type: new Abstract: Large language models (LLMs) increasingly support complex professional tasks, yet their capabilities in rule-intensive document review remain insufficie

model-releasesarxiv-cs-cl
7 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation

DGX agent

arXiv:2608.05999v1 Announce Type: new Abstract: Vision-language-action (VLA) models have demonstrated remarkable capabilities in robotic manipulation by leveraging pretrained vision-language models. H

researcharxiv-cs-ro
7 Aug 2026
Model Releases

ChronoVision: Temporal Reasoning via Latent State Reconstruction

DGX agent

arXiv:2608.05631v1 Announce Type: new Abstract: Multimodal large language models excel at passive perception but struggle with complex visual cognitive tasks requiring multi-step temporal reasoning. T

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

CNM-BERT: A Drop-In Structural Embedding for Chinese Characters via Ideographic Description Sequences

DGX agent

arXiv:2608.05167v1 Announce Type: new Abstract: Token-based encoders like BERT treat Chinese characters as atomic identifiers, ignoring their recursive orthographic structure. Consequently, models rel

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

DASH: Decoupled Adaptive Surrogate - Acquisition Harness for Automated Bayesian Optimization

DGX agent

arXiv:2608.00641v2 Announce Type: replace Abstract: Bayesian optimization (BO) relies on a surrogate model and an acquisition function, yet the most suitable choices vary across tasks and optimization

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Domain-Grounded Candidate Selection for Agentic Image Editing: A Shadow Removal Case

DGX agent

arXiv:2608.06075v1 Announce Type: cross Abstract: Commercial vision-language models are reshaping computer vision, with visual priors broad enough to rival task-specific systems. This raises a natural

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

DREAM: LLM-based Dynamic Role-playing via Event-Aware Memory Graph

DGX agent

arXiv:2608.05170v1 Announce Type: cross Abstract: Role-playing agents (RPAs) have emerged as a key application of large language models, enabling immersive and high-fidelity character simulation. Accu

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Energy-Guided Flow Matching

DGX agent

arXiv:2608.05811v1 Announce Type: new Abstract: Pixel-space generative models bypass lossy latent compression, yet necessitate joint learning of global structure and fine-grained details in a high-dim

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

EschaLabs/Qwen3.6-35B-A3B-Escha-W2 · Hugging Face

DGX agent

Hey peeps. I know you're tired of low quants giving hard to believe numbers. I'm quite skeptical too and from what I tried I'm often left with the impression that the claims fall short. So this model

model-releasesr-localllama
7 Aug 2026
Model Releases

Evidential Rule Learning for Interpretable Classification with Abstention

DGX agent

arXiv:2608.05859v1 Announce Type: cross Abstract: Interpretable classification often requires more than accurate predictions for real-life deployment: models should be transparent about the evidence b

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?

DGX agent

arXiv:2608.05747v1 Announce Type: new Abstract: Spatial intelligence is fundamental to embodied agents, yet existing benchmarks focus on local spatial perception from single or few viewpoints, overloo

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

HarnessOpt-Bench: Evaluating LLMs at Harness Optimization

DGX agent

arXiv:2608.06301v1 Announce Type: new Abstract: As LLMs are increasingly deployed within agentic systems, their capabilities depend not only on the model weights but also on the harness: the prompts,

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Hyper-ES: Effective Evolution Strategies for LLM Reasoning via Descent Direction Merging

DGX agent

arXiv:2608.05541v1 Announce Type: new Abstract: Evolution Strategy (ES) is a promising alternative to gradient-based fine-tuning for resource-constrained Large Language Model (LLM) reasoning. However,

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Innocent Panels, Hateful Stories: Evaluating and Detecting Hateful Intent in Multi-Turn Visual Story Generation

DGX agent

arXiv:2608.05210v1 Announce Type: cross Abstract: Picture books and comics have long been used to disseminate hateful narratives because they are easily understood even by children, as exemplified by

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

IPV-Bench: Benchmarking Image Protection Methods under Diverse Image-to-Video Generation Scenarios

DGX agent

arXiv:2603.26154v2 Announce Type: replace Abstract: Image-to-video (I2V) generation models can be misused to animate a single image into a convincing fake video, motivating perturbation-based image pr

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

MoCA: Implicit Social Context Analysis

DGX agent

arXiv:2608.05825v1 Announce Type: new Abstract: Human social communication, such as affection and intent, is often conveyed in highly implicit ways, where underlying meanings are expressed through ind

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

NeSy-RAG: Neuro-Symbolic RAG for Explainable Question Answering

DGX agent

arXiv:2608.06292v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves question answering by grounding large language models (LLMs) in external knowledge such as text corpora. H

model-releasesarxiv-cs-cl
7 Aug 2026
Applications

nnMIL: A generalizable multiple instance learning framework for computational pathology

DGX agent

arXiv:2511.14907v2 Announce Type: replace Abstract: Computational pathology holds substantial promise for improving diagnosis and guiding treatment decisions. Recent pathology foundation models enable

applicationsarxiv-cs-cv
7 Aug 2026
Model Releases

PoolBench: A Benchmark for Pooling Strategies in Concept Representation Evaluation for Decoder-Only LLMs

DGX agent

arXiv:2608.05162v1 Announce Type: new Abstract: Pooling is a consequential but under-examined design choice in decoder-only concept representation work: practitioners must collapse token-level hidden

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Robust Native Language Identification through Agentic Decomposition

DGX agent

arXiv:2509.16666v2 Announce Type: replace Abstract: Large language models (LLMs) often achieve high performance in native language identification (NLI) benchmarks by leveraging superficial contextual

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Schema-Guided Hierarchical Information Extraction and Semantic Evaluation Using Generative AI

DGX agent

arXiv:2608.06167v1 Announce Type: new Abstract: We present a schema-based framework for extracting complex, structured information from unstructured text documents using generative AI, followed by aut

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptation

DGX agent

arXiv:2608.05785v1 Announce Type: cross Abstract: Multilingual text embedding models are commonly adapted using a single training objective across diverse tasks, despite different tasks requiring fund

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Unifying Structured and Unstructured Data Insights with BQ Search Innovations

DGX agent

Modern enterprises possess a vast amount of unstructured data, yet they frequently encounter significant challenges in managing and extracting value from it. Historically, unlocking the insights hidde

model-releasesgoogle-cloud-ai
7 Aug 2026
Model Releases

Advancing Utility Pole and Sign Detection Through Deep Learning

DGX agent

arXiv:2608.04061v1 Announce Type: new Abstract: Utility poles are an essential part of the infrastructure used to support power distribution systems and other critical public services. Their regular i

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

An active-learning framework for real-time depth perception from monocular vision streams

DGX agent

arXiv:2608.04917v1 Announce Type: new Abstract: Biological visual systems can perceive depth from monocular vision flow, continuously integrating temporal visual cues while maintaining a balance betwe

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Anthropic will design its own hardware to power Claude

DGX agent

Anthropic is hiring a custom silicon team to design proprietary chips that will power its Claude models, while still planning a multi‑chip strategy that mixes internally designed hardware with compone

model-releasesars-technica
6 Aug 2026
Model Releases

Anyone understand what the equivalent of GPT-5.6 Instant in ChatGPT is for the OpenAI API?

DGX agent

The tweet is a question from user Simon Willison (posted on 6 Aug 2026) asking which OpenAI API model corresponds to the ChatGPT “GPT‑5.6 Instant” version. No answer or clarification is included in th

model-releasessimon-willison--x
6 Aug 2026
Local Ai

Bi-Level Reinforcement Learning Pathway for Sim-to-Real Optimality

DGX agent

arXiv:2510.17709v2 Announce Type: replace-cross Abstract: Training Reinforcement Learning (RL) policies using simulation models before deployment in real-world environments is a common strategy when r

local-aiarxiv-cs-ai
6 Aug 2026
Model Releases

BIM-Native Tokenization for Constraint-Aware Room Layout Synthesis

DGX agent

arXiv:2512.04832v3 Announce Type: replace Abstract: We present a BIM-native tokenization for room-level layout synthesis in Building Information Modeling (BIM) scenes. The core contribution is represe

model-releasesarxiv-cs-cv
6 Aug 2026
Research

COMPAS: Difficulty-Aware Joint Search for Optimizing Code Generation

DGX agent

arXiv:2608.04336v1 Announce Type: cross Abstract: Code generation systems make each LLM call with a model, a prompt, and decoding settings. However, existing optimization methods usually tune only par

researcharxiv-cs-ai
6 Aug 2026
Research

Elbow-Based MoE Routing: A Training-Free Inference Time Plugin for Expert Selection

DGX agent

arXiv:2608.04401v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models enable model scaling while maintaining low inference-time compute by activating only a subset of experts per token. Howe

researcharxiv-cs-lg
6 Aug 2026
Model Releases

ExeCRE: Execution-Consistency Guided Reliability Estimation for Self-Correcting Code Generation

DGX agent

arXiv:2608.04439v1 Announce Type: cross Abstract: Large language models (LLMs) have made notable progress in code generation, but they still struggle on challenging tasks that require sophisticated al

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

From Score Matrices to Football-Aware Match-State Simulation: An Auditable LLM Harness for Exact-Score Reranking

DGX agent

arXiv:2608.05030v1 Announce Type: new Abstract: Football score forecasting combines a strong statistical core with a difficult contextual edge. Dynamic Poisson-family models estimate team strength, ex

model-releasesarxiv-cs-ai
6 Aug 2026
Research

Fundamentals of quantum Boltzmann machine learning with visible and hidden units

DGX agent

arXiv:2512.19819v2 Announce Type: replace-cross Abstract: One of the primary applications of classical Boltzmann machines is generative modeling, wherein the goal is to tune the parameters of a model

researcharxiv-cs-lg
6 Aug 2026
Model Releases

@GoogleDeepMind Humanoid legs or wheeled rovers? Should robots be cracking eggs? Watch as the @GoogleDeepMind team behind Gemini Robotics 2 …

DGX agent

On July 30 Google AI published a video announcing **Gemini Robotics 2**, an intelligence layer developed by DeepMind that aims to bring autonomous robots closer to everyday human environments. The pos

model-releasesgoogle-ai--x
6 Aug 2026
Model Releases

Hallucinations on the Board: Tool-Augmented Evaluation of LLM Chess Commentary

DGX agent

arXiv:2608.04240v1 Announce Type: cross Abstract: Superhuman game engines in domains like chess have made expert-level evaluations easily accessible, yet they communicate what is true without the natu

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

I ported vLLM's serving stack to C++20: 66 MiB binary, no Python at inference, output checked token-for-token against vLLM

DGX agent

I'm the author, so discount the enthusiasm accordingly. This is an unaffiliated community port, not endorsed by the vLLM project, which it uses to verify its correctness. What started it: I love vLLM,

model-releasesr-localllama
6 Aug 2026
Safety

Long-term Measurements: Towards a Longitudinal Understanding of Human-AI Interactions

DGX agent

arXiv:2608.02491v2 Announce Type: replace Abstract: Language models have taken on the role of a very new type of technology, by virtue of their 'human-ness' and rapid integration into users' daily liv

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

Not Truly Multilingual: Script Consistency as a Missing Dimension in VLM Evaluation

DGX agent

arXiv:2606.17188v3 Announce Type: replace-cross Abstract: Current multilingual evaluations for Vision-Language Models (VLMs) assume a one-to-one mapping between language and orthography, overlooking b

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Predict, Then Retrieve: Cross-Instance Future-State Retrieval from Video Prefixes

DGX agent

arXiv:2608.04426v1 Announce Type: cross Abstract: We introduce Predictive State Retrieval (PSR), a task in which a model observes a short video prefix and a temporal question about an object's future

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Robust and Personalized Federated Learning for Aircraft-Engine Prognostics under Benign and Adversarial Client Heterogeneity

DGX agent

arXiv:2608.04045v1 Announce Type: cross Abstract: Federated learning (FL) enables aircraft fleet operators to jointly train remaining-useful-life (RUL) models from engine sensor telemetry without shar

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Semantic Frame Interpolation

DGX agent

arXiv:2507.05173v2 Announce Type: replace Abstract: Generating intermediate video content of varying lengths based on given first and last frames, along with text prompt information, offers significan

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Strengthening Target-Language Features: SAE-Based Steering for Multilingual Inference

DGX agent

arXiv:2608.04904v1 Announce Type: new Abstract: Multilingual large language models exhibit substantial performance differences across languages, while existing adaptation methods often require paramet

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

AS-FedBridge: Pseudo-Spike Bridge Distillation for Heterogeneous ANN-SNN Federated Learning

DGX agent

arXiv:2608.03324v1 Announce Type: new Abstract: Federated learning enables collaborative model training across distributed edge devices while strictly preserving data privacy. To facilitate practical

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Balancing Efficiency and Efficacy: Training-Free Attention-Guided Switching Between Explicit and Latent Thoughts for MLLMs

DGX agent

arXiv:2608.03450v1 Announce Type: cross Abstract: Reasoning in Multimodal Large Language Models (MLLMs) requires both fine-grained visual perception and rigorous logical deduction. Explicit text-based

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Beyond Initialization Loss: A Systematic Study of Token Embedding Initialization Strategies for LLM Vocabulary Extension

DGX agent

arXiv:2608.03494v1 Announce Type: new Abstract: Vocabulary extension is an efficient way to adapt pretrained large language models (LLMs) to new languages, but the initialization of newly added token

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Beyond Representational Similarity: Source-Conditioned Description-Length Gain for Generative Plagiarism Detection and Candidate Source Reranking

DGX agent

arXiv:2608.03859v1 Announce Type: cross Abstract: Large language models (LLMs) pose challenges to academic integrity and peer review. Yet generative plagiarism detection remains an underexplored and l

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Detecting Hallucinations and Recovering Verified Answers in Arabic Islamic Question Answering

DGX agent

arXiv:2608.03720v1 Announce Type: new Abstract: Large language models can generate fluent responses to Islamic questions while introducing factual errors that are difficult to identify. This paper pre

model-releasesarxiv-cs-cl
5 Aug 2026
← Previous
1…421422423424425…1327
Next →