AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
14 Aug 2026

RealMat: Realistic Materials with Diffusion and Reinforcement Learning

ApplicationsDGX agent

arXiv:2509.01134v2 Announce Type: replace-cross Abstract: Generative models for high-quality materials are particularly desirable to make 3D content authoring more accessible. However, the majority of

Rules or Character? Scaling Laws for AI Safety Design

Model ReleasesDGX agent

arXiv:2608.13345v1 Announce Type: new Abstract: Artificial Intelligence (AI) safety systems combine character shaping (e.g., Reinforcement Learning from Human Feedback [RLHF], Constitutional AI), whic

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding

Model ReleasesDGX agent

arXiv:2608.12748v1 Announce Type: new Abstract: Referring Expression Comprehension (REC) is commonly studied under dataset-specific fine-tuning, resulting in specialist models with limited cross-datas

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SynAct: A Reasoning-Acting Large Language Model Agent for Adaptive Synthesis Optimization

AgentsDGX agent

arXiv:2608.12751v1 Announce Type: cross Abstract: Logic synthesis transforms RTL designs into gate-level netlists, where PPA results are highly sensitive to the choice of optimization commands, making

13 Aug 2026

DexterSQL: Deep Schema Exploration and Rule-based Correction for Text-to-SQL Generation

Model ReleasesDGX agent

arXiv:2608.11889v1 Announce Type: cross Abstract: Prompting-based (extit{i}.extit{e}., non-fine-tuning) Text-to-SQL methods, where underlying large language model parameters are not changed for the ta

FrontierFinance: A Challenging Benchmark for Measuring Frontier Intelligence of Finance Agents

Model ReleasesDGX agent

arXiv:2608.11683v1 Announce Type: new Abstract: AI agents are increasingly deployed for professional investment research, yet no benchmark captures the complexity of the full investor workflow. Existi

G0.5: One Autoregressive Stream for Robot Reasoning and Action

Model ReleasesDGX agent

arXiv:2608.11739v1 Announce Type: cross Abstract: The prevailing recipe for Vision-Language-Action (VLA) models couples a pretrained VLM with a separately trained flow-matching action expert. This mak

Google launches Gemini 3.7 Flash for coding, AI agent projects

Model ReleasesDGX agent

Google LLC today launched its most capable entry-level artificial intelligence model yet. Gemini 3.7 Flash is rolling out three weeks after its predecessor. Despite the short release cycle, Google eng

Information Abundance Paradox: Long-Context Training Undermines Parametric Knowledge

Model ReleasesDGX agent

arXiv:2608.12218v1 Announce Type: cross Abstract: Large language models are increasingly trained and deployed with long contexts that span documents, code repositories, and interaction histories. This

Inverse Theory of Mind Modeling for Content Recommendation: From Web Browsing to Dynamic Intelligent Interfaces

SafetyDGX agent

arXiv:2608.11354v1 Announce Type: new Abstract: Modern recommender systems treat observed actions as reliable proxies for user preferences, yet interactions often reflect exploration or comparison rat

Measure, Don't Optimize: Forecasting Recovery in LLM Unlearning

ResearchDGX agent

arXiv:2608.11408v1 Announce Type: new Abstract: Prior white-box studies show that large language models can retain latent traces of target knowledge after unlearning, even when the knowledge is no lon

Representation Finetuning for Continual Learning

Model ReleasesDGX agent

arXiv:2603.11201v3 Announce Type: replace-cross Abstract: The world is inherently dynamic, and continual learning aims to enable models to adapt to ever-evolving data streams. While pre-trained models

Surfsvr: 2D Surface Priors as 3D Geometric Regularizers for Sparse Voxel Reconstruction

Model ReleasesDGX agent

arXiv:2608.11938v1 Announce Type: new Abstract: Sparse voxel reconstruction offers an efficient representation for high-fidelity 3D modeling, yet its geometry is commonly optimized from local photomet

The Edge-based Contiguous p-median Problem with Connections to Logistics Districting

ResearchDGX agent

arXiv:2608.11230v1 Announce Type: new Abstract: This paper introduces the edge-based contiguous p-median (ECpM) problem to partition the roads in a network into a given number of compact and contiguou

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy

Model ReleasesDGX agent

arXiv:2607.11175v2 Announce Type: replace Abstract: The growing ability of large language models and vision-language models to jointly interpret and reason over images and text is reshaping medical im

This week we launched the world's most accurate document extraction agent over real-world documents. Introducing LlamaExtract Agentic Plus …

Model ReleasesDGX agent

This week we launched the world's most accurate document extraction agent over real-world documents. Introducing LlamaExtract Agentic Plus 💫 . It is a complete document extraction model+harness engine

Try Grok 4.6

Model ReleasesDGX agent

Try Grok 4.6 Grok 4.6 wins again. 👑 Grok 4.6 takes the #1 spot on GPQA Diamond with a score of 94.9%, beating GPT-5.6, Gemini 3.1 Pro, Claude Opus 5, and every other model tested by Artificial Analysi

Using BigQuery Graphs with measures for trusted agentic workloads

AgentsDGX agent

When enterprises transition from using simple chat assistants to autonomous, agentic workloads, they quickly run into a hard truth: Agents are prone to inaccurate insights when working with directly r

12 Aug 2026

Compositional Benchmark Synthesis for Hierarchical Human Action Recognition

Model ReleasesDGX agent

arXiv:2608.10765v1 Announce Type: new Abstract: Recognizing human behavior across levels of abstraction, from atomic actions to long-horizon intentions, requires data annotated along a semantic hierar

Cross-View Sequential Visual Localization with Spatio-Temporal Context Modeling for Autonomous Driving

AgentsDGX agent

arXiv:2608.10660v1 Announce Type: cross Abstract: Continuous and reliable localization is essential for autonomous driving. Cross-view visual localization matches ground images with satellite maps, pr

DashArena: Benchmarking LLMs on Interactive Analytic Dashboard Generation

Model ReleasesDGX agent

arXiv:2608.10567v1 Announce Type: new Abstract: Analytic dashboards combine coordinated views and interactions for data exploration and decision-making. Recent models can generate them from data and n

FoR-SALE: Frame of Reference-guided Spatial Adjustment in LLM-based Diffusion Editing

SafetyDGX agent

arXiv:2509.23452v2 Announce Type: replace-cross Abstract: Current text-to-image generation models, even state-of-the-art models, exhibit a significant performance gap when spatial expressions are desc

Generating Attacks for LLMs with GFlowNets

ResearchDGX agent

arXiv:2608.10171v1 Announce Type: new Abstract: The rapid advancement of Large Language Models (LLMs) has facilitated their ubiquitous integration into various domains, leading to widespread adoption.

LiquidAI/LFM2.5-VL-3B · Hugging Face

Local AiDGX agent

LFM2.5-VL-3B is a multimodal variant of LFM2.5, a family of hybrid models designed for on-device deployment. It builds on LFM2-VL-3B with further mid- and post-training. LFM2.5-VL-3B can process both

SQuaT: Self-Supervised Knowledge Distillation via Student-Aware Quantized Teacher Features

ResearchDGX agent

arXiv:2608.10709v1 Announce Type: new Abstract: Quantization-Aware Training (QAT) enables the deployment of quantized models with minimal accuracy degradation. However, in practical scenarios, trainin

Tested Nemotron 3.5 Lightning locally on coding, Hermes Agent and agentic work

Model ReleasesDGX agent

Ran the model with quants (Q5) and MTP by bartowski with llama.cpp server. It takes ~24GB ram running on M5 Pro with 48GB at about 65t/s. On some tasks it was quite the overthinker. Overall, the quali

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?

Model ReleasesDGX agent

arXiv:2608.10875v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as personal assistants. Existing evaluations, however, mostly use short, self-contained re

Where To Look? : Causal Tracing of Vision Encoders in VLM

ResearchDGX agent

arXiv:2608.10758v1 Announce Type: new Abstract: Vision-language models can describe an image with remarkable accuracy, yet a more fundamental question remains unanswered: what visual information actua

11 Aug 2026

An Expectation-Maximization Perspective on Reinforcement Learning for LLM Reasoning

Model ReleasesDGX agent

arXiv:2504.18587v2 Announce Type: replace-cross Abstract: Reinforcement learning has emerged as a powerful approach for improving the reasoning capabilities of large language models, as demonstrated b

BASIS: Breach-Aware Selective Prompt Injection Shielding with Prefill Attention Probes

ResearchDGX agent

arXiv:2608.08027v1 Announce Type: cross Abstract: Prompt injection is a critical security threat in large language model (LLM) applications, where attackers hijack model behavior by embedding maliciou

BDH-CQ: In-Context Learning with Recurrent Latent Reasoning

Model ReleasesDGX agent

arXiv:2608.09888v1 Announce Type: cross Abstract: We introduce BDH-CQ, a reasoning model that combines in-context learning with recurrent latent reasoning. Inputs presented at inference time continuou

Build it, Break it, Repeat: Benchmarking and improving LLM-manipulated disinformation detection in social media posts

Model ReleasesDGX agent

arXiv:2608.09510v1 Announce Type: cross Abstract: Detecting machine-generated disinformation on social media is increasingly difficult as large language models (LLMs) make it easier to generate and re

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives

Model ReleasesDGX agent

arXiv:2608.08160v1 Announce Type: cross Abstract: The rapid advancement of Large Language Models (LLMs) is revolutionizing AI for Games by enabling open-ended and fluid interactive storytelling. Howev

DarwinX: Evolving Agent Harnesses Through Natural Selection

Model ReleasesDGX agent

arXiv:2608.07545v1 Announce Type: cross Abstract: An LLM agent's capability depends not only on model weights but on its harness: prompts, tools, skills, and control flow. Self-improvement loops alrea

Different Feedback, Different Updates: Selective Self-Learning from User Interactions for Large Language Models

ResearchDGX agent

arXiv:2608.09109v1 Announce Type: new Abstract: User feedback offers natural supervision for persistent LLM improvement, but a single message may support multiple behavioral changes with different sco

DiffSafeMerge: Mitigating Backdoor Inheritance in Diffusion Model Merging

ResearchDGX agent

arXiv:2608.09445v1 Announce Type: cross Abstract: Unconditional diffusion checkpoint merging assumes benign sources, yet a compromised public checkpoint can transfer a dormant backdoor while clean gen

From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition

SafetyDGX agent

arXiv:2308.04553v4 Announce Type: replace Abstract: Visual recognition models are prone to learning spurious correlations induced by a biased training set where certain conditions B (eg, Indoors) are

Hallucination-Free GUI Grounding via Regression-Free Layout-Aware Matching

Local AiDGX agent

arXiv:2608.09654v1 Announce Type: new Abstract: GUI agents are shifting from metadata-dependent large language models to purely visual multimodal large language models (MLLMs) that operate directly on

How Far Do Foundation Models Transfer to Infant Signals? A Cross-Dataset Transfer Audit with a Unified Need Ontology

ResearchDGX agent

arXiv:2608.08989v1 Announce Type: cross Abstract: Public infant cry corpora are small, label-incompatible, and almost always evaluated one corpus at a time. We ask what this practice hides and what fi

IntelliAudit: Using Large Language Models to Evaluate Audit Controls

AgentsDGX agent

arXiv:2608.07688v1 Announce Type: new Abstract: IT audits require auditors to judge whether heterogeneous organizational evidence satisfies semantic security and compliance controls. This judgment is

LHSDet: High-Resolution AI-Generated Image Detection via Visual Question Answering

ResearchDGX agent

arXiv:2608.07863v1 Announce Type: new Abstract: Driven by advances in diffusion models and autoregressive models, the fidelity and resolution of AI-generated images now rival those of real images. How

LIRA: Local Cross-Layer Information Routing for Vision-Language-Action Decoding

Model ReleasesDGX agent

arXiv:2608.07596v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models transform representations from pretrained vision-language models (VLMs) into robot actions, yet the interface that

Long SKILL Compliance as Logical Reasoning: Closure-Grounded Detection with Scaling-Guided On-Policy Distillation

Model ReleasesDGX agent

arXiv:2608.08146v1 Announce Type: new Abstract: The increasing complexity of enterprise business scenarios has promoted the widespread adoption of long SKILL documents in agent systems, posing new cha

MCIF: Multimodal Crosslingual Instruction-Following Benchmark from Scientific Talks

Model ReleasesDGX agent

arXiv:2507.19634v4 Announce Type: replace-cross Abstract: Recent advances in large language models have laid the foundation for multimodal LLMs (MLLMs), which unify text, speech, and vision within a s

Rethinking Reasoning with MDLMs: Early Exits, Post-hoc Reasoning, and Beyond

ResearchDGX agent

arXiv:2510.19990v2 Announce Type: replace Abstract: The reasoning paradigm, where language models reason before answering, has enabled breakthroughs on tasks such as mathematical problem-solving. Whil

UNMASK: Discovering and Causally Verifying Spurious Shortcuts in Text Classifiers

Model ReleasesDGX agent

arXiv:2608.09209v1 Announce Type: new Abstract: Neural language models trained on large crowdsourced corpora frequently exploit spurious surface patterns tied to target labels without true linguistic

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction

SafetyDGX agent

arXiv:2608.09448v1 Announce Type: cross Abstract: Test-time training (TTT) offers a lightweight way to adapt vision--language--action (VLA) policies from unlabeled deployment streams, but it remains d

VeinCast: Physics-Guided Dynamic Field Graphs with Graph-Conditioned Fusion for Global Medium-Range Weather Forecasting

Model ReleasesDGX agent

arXiv:2608.09286v1 Announce Type: cross Abstract: Global medium-range weather forecasting requires modeling structured yet state-dependent interactions among heterogeneous atmospheric fields. Existing

When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output

Model ReleasesDGX agent

arXiv:2503.24191v4 Announce Type: replace-cross Abstract: Content Warning: This paper may contain unsafe or harmful content generated by LLMs that may be offensive to readers. Large Language Models (L

10 Aug 2026

Achievable 253 t/s - unsloth/Muse Glimmer 30B UD-Q5_K_M on a 5090

Model ReleasesDGX agent

Benchmarked Muse Glimmer 30B on my RTX 5090 (32GB), 262k context, UD-Q5_K_M + dflash-kquant + mmproj. Workload Stock master + DFlash ngram-simple PR #26842 + DFlash Code patch 78 t/s 57 t/s 220-253 t/

An Exploratory Evaluation of LLM-Assisted Rewriting of Moderate-Complexity Financial Sentences for DisCoCat-Based Sentiment Analysis

Model ReleasesDGX agent

arXiv:2608.07439v1 Announce Type: new Abstract: Quantum natural language processing (QNLP) provides a grammar-aware framework for text modeling, and Distributional Compositional Categorical (DisCoCat)

Are Visual Place Recognition Models Recognizing Places or Conditions? Distractor-Augmented Evaluation and Condition Suppression

ResearchDGX agent

arXiv:2608.06847v1 Announce Type: cross Abstract: Long-term Visual Place Recognition (VPR) is typically evaluated by matching queries from one condition against a database from another. Crowdsourced m

Beyond 'AI Language': The case for the idiolectal nature of LLM output

ResearchDGX agent

arXiv:2608.06589v1 Announce Type: cross Abstract: While large language model outputs are frequently analysed as a collective super variety termed 'AI language,' this chapter argues that this perspecti

CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity

SafetyDGX agent

arXiv:2608.07460v1 Announce Type: cross Abstract: While post-training improves the capabilities of large language models (LLMs), it generally lowers their output diversity and creativity, negatively i

DATAREEL: Automated Data-Driven Video Story Generation with Animations

Model ReleasesDGX agent

arXiv:2604.25220v2 Announce Type: replace Abstract: Data videos combine animated visualizations with synchronized narration to communicate quantitative information and are widely used in journalism, e

Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits

Model ReleasesDGX agent

arXiv:2608.07430v1 Announce Type: cross Abstract: Diffusion Large Language Models (DLLMs) replace autoregressive next-token prediction with iterative parallel denoising, yet their internal safety mech

Do AI Personas Grow? Analyzing and Benchmarking Personality Evolution in LLM Agents After Life Events

Model ReleasesDGX agent

arXiv:2608.06485v1 Announce Type: cross Abstract: Personality-conditioned LLM agents (PC-Agents) are increasingly used in emotional support, social simulation, and role-playing, motivating the develop

From Optimal Actions to World Models: Identifiability of Transition Kernels in Discounted MDPs

SafetyDGX agent

arXiv:2608.07301v1 Announce Type: new Abstract: We study what can be recovered about the transition probabilities of a Markov decision process from optimal actions alone. This is closely related to th

Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance avai…

Model ReleasesDGX agent

Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance available. Open weights, zero data retention, U.S. & EU hosting.

Learning in Deep Networks under Dale's Constraint

Model ReleasesDGX agent

arXiv:2608.06963v1 Announce Type: new Abstract: Biologically plausible learning models aim to explain how neural circuits can implement effective learning under the constraints of real neurons. Althou

← Previous
1…269270271272273…1034
Next →