AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,598
  • Agents7,796
  • Applications5,565
  • Concepts5
  • Hardware1,944
  • Industry6,220
  • Local Ai5,134
  • Model Releases24,972
  • Research20,928
  • Safety13,838
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,598
  • Agents7,796
  • Applications5,565
  • Concepts5
  • Hardware1,944
  • Industry6,220
  • Local Ai5,134
  • Model Releases24,972
  • Research20,928
  • Safety13,838
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,598Total entries
1Added by human
91,597Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,253 results
Model Releases

Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect

DGX agent

arXiv:2605.18808v1 Announce Type: cross Abstract: We characterize a compositional architecture of literary primitives in two instruction-tuned large language models (Llama 3.1 8B-Instruct and Gemma 2

model-releasesarxiv-cs-ai
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows

DGX agent

arXiv:2605.19099v1 Announce Type: new Abstract: We introduce DecisionBench, a benchmark substrate for emergent delegation in long-horizon agentic workflows. The substrate fixes a task suite (GAIA, tau

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Depth2Pose: A Pose-Based Benchmark for Monocular Depth Estimation without Ground-Truth Depth

DGX agent

arXiv:2605.19797v1 Announce Type: new Abstract: Monocular depth estimation has improved significantly in recent years, driven by increasingly powerful models and large-scale training data. Predicted d

model-releasesarxiv-cs-cv
20 May 2026
Local Ai

Descriptive versus Regulatory Uncertainty in Bounded Predictive Systems

DGX agent

arXiv:2605.18909v1 Announce Type: new Abstract: Any system that models the world under finite representational capacity must compress; any compression entails a prior; and the prior is the system's bi

local-aiarxiv-cs-lg
20 May 2026
Model Releases

Differential-Integral Neural Operator for Long-Term Turbulence Forecasting

DGX agent

arXiv:2509.21196v3 Announce Type: replace-cross Abstract: Accurately forecasting the long-term evolution of turbulence represents a grand challenge in scientific computing and is crucial for applicati

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

DMN: A Compositional Framework for Jailbreaking Multimodal LLMs with Multi-Image Inputs

DGX agent

arXiv:2605.18915v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are vulnerable to jailbreak attacks, which can elicit harmful responses from MLLMs. Many MLLMs support multi-

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

DocQT: Improving Document Forgery Localization Robustness via Diverse JPEG Quantization Tables

DGX agent

arXiv:2605.19688v1 Announce Type: new Abstract: Document manipulation localization models achieve strong performance on public benchmarks yet fail to generalize to operational document workflows. We i

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

EventPrune: Cascaded Event-Assisted Token Pruning for Efficient First-Person Dynamic Spatial Reasoning

DGX agent

arXiv:2605.19506v1 Announce Type: new Abstract: First-person dynamic spatial reasoning requires models to track continuous motion and precise geometric structure, but the quadratic attention cost of T

model-releasesarxiv-cs-cv
20 May 2026
Safety

Feature-Space Smoothing: Certified Robustness of Deep Representations

DGX agent

arXiv:2601.16200v3 Announce Type: replace-cross Abstract: Modern deep learning models exhibit strong capabilities across diverse applications, yet remain vulnerable to malicious inputs that induce err

safetyarxiv-cs-cv
20 May 2026
Safety

Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents

DGX agent

arXiv:2605.19604v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly act inside real workspaces, where tools and skills determine whether model reasoning becomes reliable act

safetyarxiv-cs-ai
20 May 2026
Model Releases

Gemini 3.5 announce.

DGX agent

Google introduced Gemini 3.5, its latest family of models combining frontier intelligence with action capabilities, representing a major leap forward in building more capable, intelligent agents. The

model-releasesr-chatgpt
20 May 2026
Model Releases

Generalization Bounds of Surrogate Policies for Combinatorial Optimization Problems

DGX agent

arXiv:2407.17200v3 Announce Type: replace-cross Abstract: Many real-world decision problems require solving, again and again, combinatorial optimization instances drawn from a common distribution. A r

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation

DGX agent

arXiv:2605.19890v1 Announce Type: new Abstract: Test-time adaptation (TTA) enables a pre-trained model to adapt online to an unlabeled test stream under distribution shift. While most TTA research foc

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding

DGX agent

arXiv:2605.19223v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) exhibit strong performance on standard video tasks, their ability to faithfully summarize and reason over

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Hybrid-LoRA: Bridging Full Fine-Tuning and Low-Rank Adaptation for Post-Training

DGX agent

arXiv:2605.18822v1 Announce Type: cross Abstract: Post-training has become essential for adapting large language models (LLMs) to complex downstream behaviors, including instruction following, prefere

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Less Back-and-Forth: A Comparative Study of Structured Prompting

DGX agent

arXiv:2605.20149v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for open-ended tasks, but underspecified prompts can lead to low-quality answers and additional interacti

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Library Hallucinations in LLM-Generated Code: A Risk Analysis Grounded in Developer Queries

DGX agent

arXiv:2509.22202v3 Announce Type: replace-cross Abstract: Large language models (LLMs) now play a central role in code generation, yet they continue to hallucinate, frequently inventing non-existent l

model-releasesarxiv-cs-cl
20 May 2026
Research

Lossless Anti-Distillation Sampling

DGX agent

arXiv:2605.18829v1 Announce Type: new Abstract: Frontier commercial generative models face a growing threat from distillation, whereby a distiller harvests generated responses and trains a competing m

researcharxiv-cs-lg
20 May 2026
Model Releases

LWiAI Podcast #245 - TML-Interaction, Claude For Legal, Sam Altman on Stand

DGX agent

This podcast episode covers three main topics: TML-Interaction (likely a new AI model or technical development), the application of Claude AI in legal settings and use cases, and Sam Altman's testimon

model-releaseslast-week-in-ai
20 May 2026
Model Releases

MIRO: MultI-Reward cOnditioned pretraining improves T2I quality and efficiency

DGX agent

arXiv:2510.25897v2 Announce Type: replace Abstract: The default paradigm of post-training text-to-image generators includes post-hoc selection of generated images, and subsequent training with one rew

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs

DGX agent

arXiv:2511.14159v2 Announce Type: replace Abstract: Evaluating the robustness of Large Vision-Language Models (LVLMs) is essential for their continued development and responsible deployment in real-wo

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Next-Acceleration-Scale Prediction for Autoregressive MRI Reconstruction

DGX agent

arXiv:2605.19354v1 Announce Type: cross Abstract: MRI reconstruction is an inherently ill-posed inverse problem, since incomplete measurements admit many plausible solutions. This ambiguity becomes mo

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

DGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

model-releasesarxiv-cs-cv
20 May 2026
Safety

Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering

DGX agent

arXiv:2605.19220v1 Announce Type: cross Abstract: Uncertainty Quantification (UQ) is widely regarded as the primary safeguard for deploying Large Language Models (LLMs) in high-stakes domains. However

safetyarxiv-cs-ai
20 May 2026
Model Releases

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling

DGX agent

arXiv:2605.20052v1 Announce Type: cross Abstract: Automatic report labeling facilitates the identification of clinical findings from unstructured text and enables large-scale annotation for medical im

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

RECIPE: Procedural Planning via Grounding in Instructional Video

DGX agent

arXiv:2605.19976v1 Announce Type: new Abstract: Visual planning asks a model to generate the remaining steps of a procedure in natural language given a partial video context and a goal. Progress on th

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Retrieval-Augmented Generation for Natural Language Processing: A Survey

DGX agent

arXiv:2407.13193v4 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong empirical performance in various fields, benefiting from their huge amount of parameters that stor

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Smooth Piecewise Cutting for Neural Operator to Handle Discontinuities and Sharp Transitions

DGX agent

arXiv:2605.19823v1 Announce Type: cross Abstract: Neural operators have achieved strong performance in learning solution operators of partial differential equations (PDEs), but their inherently contin

model-releasesarxiv-cs-ai
20 May 2026
Applications

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence

DGX agent

arXiv:2505.23747v2 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced performance on 2D visual tasks. However, improving

applicationsarxiv-cs-ai
20 May 2026
Model Releases

Tail Annealing for Heavy-Tailed Flow Matching

DGX agent

arXiv:2605.20068v1 Announce Type: cross Abstract: Standard generative models struggle with heavy-tailed data: Lipschitz architectures cannot produce power-law tails from Gaussian noise, and interpolat

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning

DGX agent

arXiv:2605.19358v1 Announce Type: new Abstract: Entropy-based deep reasoning has emerged as a promising direction for improving the reasoning capabilities of Large Language Models (LLMs), but existing

model-releasesarxiv-cs-cl
20 May 2026
Safety

TEMPO: Temporal Enforcement via Mode-Separated Policy Optimization for Trustworthy LLM Backtesting

DGX agent

arXiv:2605.18843v1 Announce Type: new Abstract: Backtesting large language models on historical events requires reasoning exclusively from information available before a specified cutoff date. Yet mod

safetyarxiv-cs-lg
20 May 2026
Model Releases

The Annotation Scarcity Paradox in Low-Resource NLP Evaluation: A Decade of Acceleration and Emerging Constraints

DGX agent

arXiv:2605.19066v1 Announce Type: new Abstract: Over the past decade, low-resource natural language processing (NLP) has experienced explosive growth, propelled by cross-lingual transfer, massively mu

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measur…

DGX agent

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measuring every biomarker, or @sytses openly sharing and analyzing

model-releasesclem-delangue--x
20 May 2026
Model Releases

The Silent Hyperparameter: Quantifying the Impact of Inference Backends on LLM Reproducibility

DGX agent

arXiv:2605.19537v1 Announce Type: new Abstract: Progress in LLMs is increasingly measured through standardized benchmarks, where state-of-the-art improvements are often separated by fractions of a per

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Towards Consistent Detection of Cognitive Distortions: LLM-Based Annotation and Dataset-Agnostic Evaluation

DGX agent

arXiv:2511.01482v2 Announce Type: replace Abstract: Text-based automated Cognitive Distortion detection is a challenging task due to its subjective nature, with low agreement scores observed even amon

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

DGX agent

arXiv:2605.18859v1 Announce Type: cross Abstract: LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single user reque

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

What Do Evolutionary Coding Agents Evolve?

DGX agent

arXiv:2605.20086v1 Announce Type: cross Abstract: Recent work pairs LLMs with evolutionary search to iteratively generate, modify, and select code using task-specific feedback. These systems have prod

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection

DGX agent

arXiv:2601.22569v2 Announce Type: replace-cross Abstract: Large language model (LLM) based agents are increasingly used to automate financial transactions, yet their reliance on contextual reasoning e

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Feature-Driven Framework for Software Fault Prediction

DGX agent

arXiv:2605.17611v1 Announce Type: cross Abstract: Software fault prediction (SFP) is a critical task in software engineering, enabling early identification of faults in modules to improve software qua

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

A Machine With Human-Like Memory Systems

DGX agent

arXiv:2204.01611v3 Announce Type: replace Abstract: Inspired by the cognitive science theory, we explicitly model an agent with both semantic and episodic memory systems, and show that it is better th

model-releasesarxiv-cs-ai
19 May 2026
Research

A Theory of Training Profit-Optimal LLMs

DGX agent

arXiv:2605.16430v1 Announce Type: cross Abstract: Scaling LLMs requires tremendous computational resources, and recent advances in AI have gone hand in hand with massive amounts of capital expenditure

researcharxiv-cs-ai
19 May 2026
Local Ai

AdaptiveLoad: Towards Efficient Video Diffusion Transformer Training

DGX agent

arXiv:2605.17923v1 Announce Type: cross Abstract: In video generation models, particularly world models, training large-scale video diffusion Transformers (such as DiT and MMDiT) poses significant com

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Advancing Narrative Long Video Generation via Training-Free Identity-Aware Memory

DGX agent

arXiv:2605.18733v1 Announce Type: new Abstract: Autoregressive video generation has improved rapidly in visual fidelity and interactivity, but it still suffers from long-term inconsistency and memory

model-releasesarxiv-cs-cv
19 May 2026
Agents

AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent

DGX agent

arXiv:2602.03955v2 Announce Type: replace Abstract: While large language model (LLM) multi-agent systems achieve superior reasoning performance through iterative debate, practical deployment is limite

agentsarxiv-cs-ai
19 May 2026
Model Releases

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech

DGX agent

arXiv:2605.17583v1 Announce Type: new Abstract: While existing text-to-speech (TTS) models exhibit high expressiveness, fine-grained control over composite instructions remains challenging due to the

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

AgentWall: A Runtime Safety Layer for Local AI Agents

DGX agent

arXiv:2605.16265v1 Announce Type: new Abstract: The safety of autonomous AI agents is increasingly recognized as a critical open problem. As agents transition from passive text generators to active ac

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing

DGX agent

arXiv:2603.23069v2 Announce Type: replace-cross Abstract: The task of authorship style transfer involves rewriting text in the style of a target author while preserving the meaning of the original tex

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…538539540541542…1381
Next →