AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlog
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
Model Releases

What Google I/O '26 means for developing agents on Google Cloud

DGX agent

At Google I/O, we introduced a unified development toolkit featuring Antigravity 2.0 and the Managed Agents API, giving developers better ways to build locally and deploy securely to the cloud on a sh

model-releasesgoogle-cloud-ai
19 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

what if we mapped older distributed systems patterns (like actor models or reactive, state-driven blackboards) to LLM agents?

DGX agent

This post explores conceptual parallels between classical distributed systems design patterns—such as actor models and reactive blackboard architectures—and their potential application to LLM agent de

agentsyohei-nakajima--x
19 May 2026
Industry

⚡️After weeks of hard work, we're thrilled to fully open-source HY World 2.0 today -- full inference code and all models! Build, explore, an…

DGX agent

⚡️After weeks of hard work, we're thrilled to fully open-source HY World 2.0 today -- full inference code and all models! Build, explore, and create your own interactive worlds with us.👇 https://githu

industryemad-mostaque--x
18 May 2026
Research

Autoguided Online Data Curation for Diffusion Model Training

DGX agent

arXiv:2509.15267v2 Announce Type: replace-cross Abstract: The costs of generative model compute rekindled promises and hopes for efficient data curation. In this work, we investigate whether recently

researcharxiv-cs-ai
18 May 2026
Research

Generative Long-term User Interest Modeling for Click-Through Rate Prediction

DGX agent

arXiv:2605.15905v1 Announce Type: cross Abstract: Modeling long-term user interests with massive historical user behaviors enhances click-through rate (CTR) prediction performance in advertising and r

researcharxiv-cs-ai
18 May 2026
Tools

Introducing Composer 2.5, our most powerful model yet. It's more intelligent, better at sustained work on long-running tasks, and more relia…

DGX agent

Introducing Composer 2.5, our most powerful model yet. It's more intelligent, better at sustained work on long-running tasks, and more reliable at following complex instructions. For the next week, we

toolscursor--x
18 May 2026
Applications

Its a consistent theory-of-mind failure in models that are otherwise suprisingly good at theory-of-mind

DGX agent

Large language models demonstrate a specific and persistent theory-of-mind failure despite excelling at other theory-of-mind tasks, suggesting a particular gap in their reasoning about beliefs, intent

applicationsethan-mollick--x
18 May 2026
Safety

Monotone and Separable Set Functions: Characterizations and Neural Models

DGX agent

arXiv:2510.23634v3 Announce Type: replace-cross Abstract: Motivated by applications for set containment problems, we consider the following fundamental problem: can we design set-to-vector functions s

safetyarxiv-cs-ai
18 May 2026
Safety

Note that as of last summer Waymo said that LLMs/VLMs were experimental, and that their foundation model then could “process only a small am…

DGX agent

Note that as of last summer Waymo said that LLMs/VLMs were experimental, and that their foundation model then could “process only a small amount of image frames, does not incorporate accurate 3D sensi

safetygary-marcus--x
18 May 2026
Research

One Pass Is Not Enough: Recursive Latent Refinement for Generative Models

DGX agent

arXiv:2605.15309v1 Announce Type: new Abstract: Despite remarkable progress, image generation is far from solved. The dominant metric, FID, conflates sample fidelity with mode coverage and is close to

researcharxiv-cs-cv
18 May 2026
Hardware

Together with SpaceXAI, we’re training a significantly larger model from scratch, using 10x more total compute. With Colossus 2’s million H1…

DGX agent

Together with SpaceXAI, we’re training a significantly larger model from scratch, using 10x more total compute. With Colossus 2’s million H100-equivalents and our combined data and training techniques

hardwarecursor--x
18 May 2026
Model Releases

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation

DGX agent

arXiv:2605.16079v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have shown significant progress in video understanding, yet they face substantial challenges in tasks requiring p

model-releasesarxiv-cs-ai
18 May 2026
Agents

The best feature of @xai Grok Build right now is how it handles subagents and personas. Most people still treat the model like one very smar…

DGX agent

The best feature of @xai Grok Build right now is how it handles subagents and personas. Most people still treat the model like one very smart intern that has to do everything at once. Grok Build took

agentselon-musk--x
17 May 2026
Agents

Interesting interpretability paper on tool-using agents. The authors probe hidden states and find the model often recognizes it should call …

DGX agent

Interesting interpretability paper on tool-using agents. The authors probe hidden states and find the model often recognizes it should call a tool, but fails to actually call one. The mismatch ranges

agentsdair-ai--x
16 May 2026
Research

A Hardware-Aware, Per-Layer Methodology for Post-Training Quantization of Large Language Models

DGX agent

arXiv:2605.14929v1 Announce Type: new Abstract: Scaled Outer Product (SOP) is a post-training quantization methodology for large language model weights, designed to deliver near-lossless fidelity at 4

researcharxiv-cs-lg
15 May 2026
Agents

AI Knows When It's Being Watched: Functional Strategic Action and Contextual Register Modulation in Large Language Models

DGX agent

arXiv:2605.15034v1 Announce Type: cross Abstract: Large language models (LLMs) have been extensively studied from computational and cognitive perspectives, yet their behavior as communicative actors i

agentsarxiv-cs-ai
15 May 2026
Research

Are Candidate Models Really Needed for Active Learning?

DGX agent

arXiv:2605.14689v1 Announce Type: new Abstract: Deep learning has profoundly impacted domains such as computer vision and natural language processing by uncovering complex patterns in vast datasets. H

researcharxiv-cs-cv
15 May 2026
Research

Artificial Intelligence-Assistant Cardiotocography: Unified Model for Signal Reconstruction, Fetal Heart Rate Analysis, and Variability Assessment

DGX agent

arXiv:2605.14242v1 Announce Type: cross Abstract: The monitoring of fetal heart rate (FHR) and the assessment of its variability are crucial for preventing fetal compromise and adverse outcomes. Howev

researcharxiv-cs-ai
15 May 2026
Safety

AutoMoT: A Unified Vision-Language-Action Model with Asynchronous Mixture-of-Transformers for End-to-End Autonomous Driving

DGX agent

arXiv:2603.14851v3 Announce Type: replace Abstract: Integrating vision-language models (VLMs) into end-to-end (E2E) autonomous driving (AD) systems has shown promise in improving scene understanding.

safetyarxiv-cs-cv
15 May 2026
Tutorials

Causal Time Series Generation via Diffusion Models

DGX agent

arXiv:2509.20846v3 Announce Type: replace Abstract: Time series generation (TSG) synthesizes realistic sequences and has achieved remarkable success. Among TSG, conditional models generate sequences g

tutorialsarxiv-cs-lg
15 May 2026
Research

Covariance-aware sampling for Diffusion Models

DGX agent

arXiv:2605.13910v1 Announce Type: cross Abstract: We present a covariance-aware sampler that improves the quality of pixel-space Diffusion Model (DM) sampling in the few-step regime. We hypothesize th

researcharxiv-cs-cv
15 May 2026
Safety

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models

DGX agent

arXiv:2605.15055v1 Announce Type: cross Abstract: Reinforcement learning has emerged as a powerful tool for improving diffusion-based text-to-image models, but existing methods are largely limited to

safetyarxiv-cs-cv
15 May 2026
Safety

Dimension-Level Intent Fidelity Evaluation for Large Language Models: Evidence from Structured Prompt Ablation

DGX agent

arXiv:2605.14517v1 Announce Type: cross Abstract: Holistic evaluation scores capture overall output quality but do not distinguish whether a model reproduced the structural form of a user's request fr

safetyarxiv-cs-ai
15 May 2026
Safety

Distribution Corrected Offline Data Distillation for Large Language Models

DGX agent

arXiv:2605.14071v1 Announce Type: new Abstract: Distilling reasoning traces from strong large language models into smaller ones is a promising route to improve intelligence in resource-constrained set

safetyarxiv-cs-cl
15 May 2026
Research

Factorization-Error-Free Discrete Diffusion Language Model via Speculative Decoding

DGX agent

arXiv:2605.14305v1 Announce Type: new Abstract: Discrete diffusion language models improve generation efficiency through parallel token prediction, but standard X_0 prediction methods introduce factor

researcharxiv-cs-cl
15 May 2026
Research

Geometry-Aware Decoding with Wasserstein-Regularized Truncation and Mass Penalties for Large Language Models

DGX agent

arXiv:2602.10346v2 Announce Type: replace Abstract: Large language models (LLMs) must balance diversity and creativity against logical coherence in open-ended generation. Existing truncation-based sam

researcharxiv-cs-cl
15 May 2026
Model Releases

IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation

DGX agent

arXiv:2605.14712v1 Announce Type: cross Abstract: Robot imitation data are often multimodal: similar visual-language observations may be followed by different action chunks because human demonstrators

model-releasesarxiv-cs-ai
15 May 2026
Tutorials

Large Language Models for Web Accessibility: A Systematic Literature Review

DGX agent

arXiv:2605.13873v1 Announce Type: cross Abstract: Web accessibility aims to ensure that web content and services are usable by people with diverse abilities. In recent years, Large Language Models (LL

tutorialsarxiv-cs-ai
15 May 2026
Model Releases

Learning Cross-Coupled and Regime Dependent Dynamics for Aerial Manipulation

DGX agent

arXiv:2605.14805v1 Announce Type: new Abstract: Accurate dynamics models are critical for aerial manipulators operating under complex tasks such as payload transport. However, modeling these systems r

model-releasesarxiv-cs-ro
15 May 2026
Model Releases

LLMs Should Express Uncertainty Explicitly

DGX agent

arXiv:2604.05306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often produce confident yet incorrect answers, which can lead to risky failures in real-world applications. We st

model-releasesarxiv-cs-ai
15 May 2026
Research

MPU: Towards Secure and Privacy-Preserving Knowledge Unlearning for Large Language Models

DGX agent

arXiv:2602.23798v2 Announce Type: replace-cross Abstract: Machine unlearning for large language models often faces a privacy dilemma in which strict constraints prohibit sharing either the server's pa

researcharxiv-cs-ai
15 May 2026
Industry

New Template Batch Drop SELFIE modeled by @paranoidream Classic Bedroom Mirror https://grok.com/imagine/templates/8a017403-bbad-428c-8900-5e…

DGX agent

New Template Batch Drop SELFIE modeled by @paranoidream Classic Bedroom Mirror https://grok.com/imagine/templates/8a017403-bbad-428c-8900-5e338c7f3eb4 Post-Shower Fresh Look https://grok.com/imagine/t

industryelon-musk--x
15 May 2026
Research

Non-linear Interventions on Large Language Models

DGX agent

arXiv:2605.14749v1 Announce Type: cross Abstract: Intervention is one of the most representative and widely used methods for understanding the internal representations of large language models (LLMs).

researcharxiv-cs-ai
15 May 2026
Local Ai

Ollama now supports Codex app! To try it, update to the latest Ollama 0.24, and run: ollama launch codex-app Select an open model to use wit…

DGX agent

Ollama 0.24 now supports integration with a Codex app, allowing users to run open-source models through the application after updating to the latest version. The feature enables users to select and ut

local-aiollama--x
15 May 2026
Safety

Persian MusicGen: A Large-Scale Dataset and Culturally-Aware Generative Model for Persian Music

DGX agent

arXiv:2605.14765v1 Announce Type: cross Abstract: Persian music, with its unique tonalities, modal systems (Dastgah), and rhythmic structures, presents significant challenges for music generation mode

safetyarxiv-cs-cl
15 May 2026
Applications

ReasonCache: Accelerating Large Reasoning Model Serving through KV Cache Sharing

DGX agent

arXiv:2507.21433v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) are becoming integral to many AI inference systems, enhancing their capabilities with advanced reasoning. Howeve

applicationsarxiv-cs-ai
15 May 2026
Safety

Reasoning Model Is Superior LLM-Judge, Yet Suffers from Biases

DGX agent

arXiv:2601.03630v2 Announce Type: replace Abstract: This paper presents the first systematic comparison investigating whether Large Reasoning Models (LRMs) are superior judges to non-reasoning LLMs. O

safetyarxiv-cs-cl
15 May 2026
Research

Scalable Subset Selection in Linear Mixed Models

DGX agent

arXiv:2506.20425v3 Announce Type: replace-cross Abstract: Linear mixed models (LMMs), which incorporate fixed and random effects, are key tools for analyzing heterogeneous data, such as in personalize

researcharxiv-cs-lg
15 May 2026
Tutorials

SeesawNet: Towards Non-stationary Time Series Forecasting with Balanced Modeling of Common and Specific Dependencies

DGX agent

arXiv:2605.14551v1 Announce Type: new Abstract: Instance normalization (IN) is widely used in non-stationary multivariate time series forecasting to reduce distribution shifts and highlight common pat

tutorialsarxiv-cs-lg
15 May 2026
Research

Separating Intrinsic Ambiguity from Estimation Uncertainty in Deep Generative Models for Linear Inverse Problems

DGX agent

arXiv:2605.15050v1 Announce Type: new Abstract: Recently, deep generative models have been used for posterior inference in inverse problems, including high-stakes applications in medical imaging and s

researcharxiv-cs-lg
15 May 2026
Model Releases

Thinking Ahead: Prospection-Guided Retrieval of Memory with Language Models

DGX agent

arXiv:2605.14177v1 Announce Type: cross Abstract: Long-horizon personalization requires dialogue assistants to retrieve user-specific facts from extended interaction histories. In practice, many relev

model-releasesarxiv-cs-ai
15 May 2026
Safety

Training ML Models with Predictable Failures

DGX agent

arXiv:2605.15134v1 Announce Type: new Abstract: Estimating how often an ML model will fail at deployment scale is central to pre-deployment safety assessment, but a feasible evaluation set is rarely l

safetyarxiv-cs-lg
15 May 2026
Research

Where Should Diffusion Enter a Language Model? Geometry-Guided Hidden-State Replacement

DGX agent

arXiv:2605.14368v1 Announce Type: cross Abstract: Continuous diffusion language models lag behind autoregressive transformers, partly because diffusion is applied in spaces poorly suited to language d

researcharxiv-cs-ai
15 May 2026
Safety

AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation

DGX agent

arXiv:2605.13724v1 Announce Type: cross Abstract: Few-step video generation has been significantly advanced by consistency distillation. However, the performance of consistency-distilled models often

safetyarxiv-cs-ai
14 May 2026
Research

Bridging the Missing-Modality Gap: Improving Text-Only Calibration of Vision Language Models

DGX agent

arXiv:2605.12517v1 Announce Type: cross Abstract: Vision-language models (VLMs) are often deployed on text-only inputs, although they are trained with images. We find that removing the vision modality

researcharxiv-cs-ai
14 May 2026
Local Ai

Comparing tokens per second of common models

DGX agent

This Reddit post likely compares inference performance metrics across popular language models running on Ollama, measuring tokens per second as a key performance indicator. The post would help users a

local-air-ollama
14 May 2026
Research

Debunking Grad-ECLIP: A Comprehensive Study on Its Incorrectness and Fundamental Principles for Model Interpretation

DGX agent

arXiv:2605.12952v1 Announce Type: new Abstract: Grad-ECLIP is published at ICML 2024 and represents a new Transformer interpretation technical route (intermediate features-based). First, this paper de

researcharxiv-cs-cv
14 May 2026
Agents

DID SOMEONE SAY FREE TOKENS IN FLEET???? Yes it's true. Fleet now has a built in model powered by @FireworksAI_HQ that's free for all Develo…

DGX agent

DID SOMEONE SAY FREE TOKENS IN FLEET???? Yes it's true. Fleet now has a built in model powered by @FireworksAI_HQ that's free for all Developer & Plus plan users. Try it out today (or before @hwchase1

agentsharrison-chase--x
14 May 2026
← Previous
1…234235236237238…1272
Next →