AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
Applications

SpectCount: Spectrotemporal Counting via Synthetic Signals Improves Large Audio Language Models

DGX agent

arXiv:2606.06907v1 Announce Type: cross Abstract: Large audio language models (LALMs) extend large language models with an audio encoder and large-scale audio data. However, the scarcity of high-quali

applicationsarxiv-cs-ai
8 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models

DGX agent

arXiv:2602.02600v3 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have recently emerged as a competitive alternative to autoregressive (AR) models, offering parallel decoding,

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

Textual Supervision Enhances Geospatial Representations in Vision-Language Models

DGX agent

arXiv:2606.07172v1 Announce Type: cross Abstract: Geospatial understanding is a critical yet underexplored dimension in the development of machine learning systems for tasks such as image geolocation

model-releasesarxiv-cs-ai
8 Jun 2026
Research

The Dual Mechanisms of Spatial Variable Binding in Vision-Language Models

DGX agent

arXiv:2603.22278v2 Announce Type: replace Abstract: Many multimodal tasks, such as image captioning and visual question answering, require vision-language models (VLMs) to bind objects with their prop

researcharxiv-cs-cv
8 Jun 2026
Hardware

Xiaomi just claimed 1,000+ tps on a 1T model using a standard 8-GPU server

DGX agent

Xiaomi achieved over 1,000 tokens per second output from a 1 trillion-parameter model using a single standard 8-GPU commodity node through extreme model-system codesign . The approach combines FP4 qua

hardwarer-localllama
8 Jun 2026
Agents

A Taxonomy of Runtime Faults in Model Context Protocol Servers

DGX agent

arXiv:2606.05339v1 Announce Type: cross Abstract: MCP (Model Context Protocol) enables LLMs (Large Language Models) to interact with external tools and data sources via a standardized protocol. Its ra

agentsarxiv-cs-ai
6 Jun 2026
Safety

An Infectious Disease Spread Simulation Based on Large Language Model Decision Making

DGX agent

arXiv:2606.06360v1 Announce Type: new Abstract: Modelling individual decision-making during infectious disease outbreaks is crucial for understanding behavioural dynamics and informing effective publi

safetyarxiv-cs-ai
6 Jun 2026
Applications

CausalPOI: Spatio-Temporal Graph-Based Causal Modeling for Cold-Start POI Check-in Forecasting

DGX agent

arXiv:2606.05413v1 Announce Type: cross Abstract: As urban environments continue to evolve rapidly, accurately modeling the dynamic behaviour of Points of Interest is essential for supporting data-dri

applicationsarxiv-cs-ai
6 Jun 2026
Tutorials

EEGDancer: Dynamic Emotion Latent Space Masked Modeling with Reinforcement Learning for EEG Continuous Emotion Prediction

DGX agent

arXiv:2606.05855v1 Announce Type: cross Abstract: Continuous electroencephalography (EEG) emotion prediction aims to model the temporal evolution of human emotional states from EEG signals. Unlike con

tutorialsarxiv-cs-ai
6 Jun 2026
Tutorials

In-Training Defenses against Emergent Misalignment in Language Models

DGX agent

arXiv:2508.06249v3 Announce Type: replace-cross Abstract: Fine-tuning lets practitioners repurpose aligned large language models (LLMs) for new domains, yet recent work reveals emergent misalignment (

tutorialsarxiv-cs-ai
6 Jun 2026
Research

Pattern Selectivity is Not Task-Causal Structure: A Cross-Architecture Mechanistic Study of Composed-Task Circuits in 1B-Class Language Models

DGX agent

arXiv:2606.05378v1 Announce Type: cross Abstract: We test whether a single screen-and-ablate recipe -- identify attention-head circuits by task-pattern selectivity, then verify by causal ablation agai

researcharxiv-cs-ai
6 Jun 2026
Applications

RAG Security and Privacy: Formalizing the Threat Model and Attack Surface

DGX agent

arXiv:2509.20324v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) is an emerging approach in natural language processing that combines large language models (LLMs) with ex

applicationsarxiv-cs-ai
6 Jun 2026
Agents

RAINO: Anchoring Agents in Reality, A Systematic Review and Conceptual Framework for Realism in Agent-Based Modelling

DGX agent

arXiv:2606.05167v1 Announce Type: cross Abstract: Realism is a central yet seemingly under-theorized concept in Agent-Based Modelling. This paper presents a Systematic Literature Review, aiming to ide

agentsarxiv-cs-ai
6 Jun 2026
Model Releases

A research team that includes Huawei says it successfully used Huawei's Ascend 910C chips for DeepSeek V4 Pro model's post-training, amid increased US sanctions (Coco Feng/South China Morning Post)

DGX agent

Coco Feng / South China Morning Post: A research team that includes Huawei says it successfully used Huawei's Ascend 910C chips for DeepSeek V4 Pro model's post-training, amid increased US sanctions —

model-releasestechmeme
5 Jun 2026
Research

AdaPLD: Adaptive Retrieval and Reuse for Efficient Model-Free Speculative Decoding

DGX agent

arXiv:2606.05742v1 Announce Type: new Abstract: Speculative decoding accelerates generation by verifying multiple drafted tokens in a single target-model forward pass, reducing sequential decoding ite

researcharxiv-cs-cl
5 Jun 2026
Applications

Also, a lot depends on Chinese labs continuing to ship open weights models. If they stop, the frontier falls further and further behind to t…

DGX agent

Also, a lot depends on Chinese labs continuing to ship open weights models. If they stop, the frontier falls further and further behind to those who want to use local/fine-tuned models. I think this i

applicationsethan-mollick--x
5 Jun 2026
Model Releases

Channel-Wise Mixed-Precision Quantization for Large Language Models

DGX agent

arXiv:2410.13056v4 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable success across a wide range of language tasks, but their deployment on edge devices remain

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Localizing Prompt Ambiguity in Large Language Models with Probe-Targeted Attribution

DGX agent

arXiv:2606.05486v1 Announce Type: new Abstract: Prompt ambiguity is a common source of failure in large language models, but is difficult to localize because it is a latent property of the prompt, whi

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Sources say xAI used Claude models for distillation and training, including using personal accounts and the intermediary service Blackbox AI after being cut off (Grace Kay/The Information)

DGX agent

Grace Kay / The Information: Sources say xAI used Claude models for distillation and training, including using personal accounts and the intermediary service Blackbox AI after being cut off — SpaceX's

model-releasestechmeme
5 Jun 2026
Research

The Invisible Hand of Physics: When Video Diffusion Models Know More Than They Show

DGX agent

arXiv:2606.05328v1 Announce Type: cross Abstract: Modern video diffusion models generate increasingly realistic and temporally coherent videos, motivating their use as candidate world simulators. Yet

researcharxiv-cs-cv
5 Jun 2026
Safety

VOLD: Reasoning Transfer from LLMs to Vision-Language Models via On-Policy Distillation

DGX agent

arXiv:2510.23497v3 Announce Type: replace Abstract: Training vision-language models (VLMs) for complex reasoning remains a challenging task, i.a. due to the scarcity of high-quality image-text reasoni

safetyarxiv-cs-cv
5 Jun 2026
Model Releases

AdaKoop: Efficient Modeling of Nonlinear Dynamics from Nonstationary Data Streams with Koopman Operator Regression

DGX agent

arXiv:2606.04930v1 Announce Type: cross Abstract: Real-time data analysis requires the ability to accurately and adaptively address nonlinear dynamics in a nonstationary data stream while preserving c

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Automatic Generation of Titles for Research Papers Using Language Models

DGX agent

arXiv:2606.05085v1 Announce Type: cross Abstract: The title of a research paper conveys its primary idea and, occasionally, its conclusions in a clear and concise manner. Choosing an appropriate title

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

BreastGPT: A Multimodal Large Language Model for the Full Spectrum of Breast Cancer Clinical Routine

DGX agent

arXiv:2606.04911v1 Announce Type: cross Abstract: Breast cancer remains a leading cause of cancer-related mortality among women. Its clinical management requires multimodal reasoning across a clinical

model-releasesarxiv-cs-cl
4 Jun 2026
Research

CLAW: Learning Continuous Latent Action World Models via Adversarial Latent Regularization

DGX agent

arXiv:2606.04130v1 Announce Type: new Abstract: We introduce CLAW, a fully end-to-end self-supervised framework for learning a world model jointly with continuous latent action representations directl

researcharxiv-cs-ro
4 Jun 2026
Model Releases

Evaluating Zero-Shot and One-Shot Adaptation of Small Language Models in Leader-Follower Interaction

DGX agent

arXiv:2602.23312v3 Announce Type: replace-cross Abstract: Leader-follower interaction is an important paradigm in human-robot interaction (HRI). Yet, assigning roles in real time remains challenging f

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

EvoPrompt: Guided Prompt Evolution for Vision-Language Models Adaptation

DGX agent

arXiv:2603.09493v2 Announce Type: replace-cross Abstract: The adaptation of large-scale vision-language models (VLMs) to downstream tasks with limited labeled data remains a significant challenge. Whi

model-releasesarxiv-cs-ai
4 Jun 2026
Tutorials

Geometry-Preserving Encoder/Decoder in Latent Generative Models

DGX agent

arXiv:2501.09876v3 Announce Type: replace-cross Abstract: Generative modeling aims to generate new data samples that resemble a given dataset. When using diffusion models for this task, one of the mai

tutorialsarxiv-cs-lg
4 Jun 2026
Model Releases

Introducing NVIDIA Nemotron 3 Ultra. A frontier smart open model built for long-running agents that need to plan, reason, use tools and keep…

DGX agent

Introducing NVIDIA Nemotron 3 Ultra. A frontier smart open model built for long-running agents that need to plan, reason, use tools and keep working across complex coding, research and enterprise work

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-l…

DGX agent

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-latency multilingual speech recognition. AI natives can now b

model-releasestogether-ai--x
4 Jun 2026
Research

Learning Control-Affine Reduced-Order Models via Autoencoders

DGX agent

arXiv:2606.05045v1 Announce Type: cross Abstract: We present in this paper a framework for the identification of control-affine reduced-order models (ROMs). The proposed method utilizes autoencoders (

researcharxiv-cs-lg
4 Jun 2026
Model Releases

Learning Long Range Spatio-Temporal Representations over Continuous Time Dynamic Graphs with State Space Models

DGX agent

arXiv:2606.04672v1 Announce Type: cross Abstract: Continuous-time dynamic graphs (CTDGs) provide a richer framework to capture fine-grained temporal patterns in evolving relational data. Long-range in

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world infere…

DGX agent

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world inference efficiency for long-context agentic workloads. Everythin

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

New Benchmarking Shows Limited Generalization Power of TCR Antigenic Epitope Prediction Models

DGX agent

arXiv:2606.04994v1 Announce Type: new Abstract: Accurate computational prediction of T cell receptor (TCR) antigen specificity would transform the study of T cell biology and enable scalable immune en

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running …

DGX agent

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running autonomous agents. Think use cases like coding agents, deep

model-releasesfireworks-ai--x
4 Jun 2026
Model Releases

Proof-Carrying Agent Actions: Model-Agnostic Runtime Governance for Heterogeneous Agent Systems

DGX agent

arXiv:2606.04104v1 Announce Type: cross Abstract: Agent systems execute through runtimes with very different control points: local coding tools, framework SDKs, managed agent platforms, API gateways,

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Pulled the trigger today and switched 100% of Lindy traffic to DeepSeek v4, churning from Anthropic models. Saves us millions of $ and we're…

DGX agent

Pulled the trigger today and switched 100% of Lindy traffic to DeepSeek v4, churning from Anthropic models. Saves us millions of $ and we're actually seeing an *increase* in performance on many core u

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

Robotics startup Generalist, which released its GEN-1 model to complete short physical tasks in April, raised 400M led by Radical Ventures at a 2B valuation (Dina Bass/Bloomberg)

DGX agent

Dina Bass / Bloomberg: Robotics startup Generalist, which released its GEN-1 model to complete short physical tasks in April, raised 400M led by Radical Ventures at a 2B valuation — The company raised

model-releasestechmeme
4 Jun 2026
Research

SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning Verification

DGX agent

arXiv:2606.04579v1 Announce Type: new Abstract: While Process Reward Models (PRMs) have achieved remarkable success in mathematical reasoning, their application in complex scientific domains-such as b

researcharxiv-cs-ai
4 Jun 2026
Model Releases

That's a badass title, and it's true! Every day, it gets harder and harder to create tests that AI models can't beat. Reality is humanity's …

DGX agent

That's a badass title, and it's true! Every day, it gets harder and harder to create tests that AI models can't beat. Reality is humanity's real last exam. Andon Labs' Real-World AI Evals: Claude call

model-releasesswyx--x
4 Jun 2026
Safety

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models

DGX agent

arXiv:2602.19101v2 Announce Type: replace-cross Abstract: Value alignment of Large Language Models (LLMs) requires us to empirically measure these models' actual, acquired representation of value. Amo

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Alibaba releases Qwen3.7-Plus, a multimodal proprietary model with a 1M-token context window, costing $2 per 1M tokens, 60% less than text-only Qwen3.7-Max (Carl Franzen/VentureBeat)

DGX agent

Carl Franzen / VentureBeat: Alibaba releases Qwen3.7-Plus, a multimodal proprietary model with a 1M-token context window, costing $2 per 1M tokens, 60% less than text-only Qwen3.7-Max — However, like

model-releasestechmeme
3 Jun 2026
Hardware

Announcing Foundry Managed Compute: Run open models in Microsoft Foundry

DGX agent

Microsoft Foundry Managed Compute is a new GPU platform-as-a-service for hosting open-source and custom AI models behind the same endpoint, SDKs, and bill as frontier models. The post Announcing Found

hardwaremicrosoft-foundry
3 Jun 2026
Applications

AugMask: Training Diffusion Models on Incomplete Tabular Data via Stochastic Augmentation and Masking

DGX agent

arXiv:2606.03347v1 Announce Type: cross Abstract: Score-based diffusion models have emerged as prominent deep generative models; however, their application to tabular data remains challenging because

applicationsarxiv-cs-ai
3 Jun 2026
Research

Bridging Auxiliary Constraints to Resolve Instruction Following in Large Reasoning Models

DGX agent

arXiv:2606.03624v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated impressive capabilities in many tasks, yet they struggle with reliably following multiple instructions,

researcharxiv-cs-ai
3 Jun 2026
Model Releases

ClinicalMC: A Benchmark for Multi-Course Clinical Decision-Making with Large Language Models

DGX agent

arXiv:2606.03157v1 Announce Type: new Abstract: Large language models (LLMs) have been widely adopted in healthcare, yet they still encounter significant challenges in complex clinical decision-making

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

DyaPlex: Full-Duplex Speech-Motion Model for Dyadic Interaction

DGX agent

arXiv:2606.03874v1 Announce Type: new Abstract: We present DyaPlex, a streaming, full-duplex speech-and-motion model designed for dyadic interaction. To capture the continuous and reciprocal nature of

agentsarxiv-cs-cv
3 Jun 2026
Local Ai

Expert-Aware Causal Tracing of Factual Recall in Sparse MoE Language Models

DGX agent

arXiv:2606.03780v1 Announce Type: new Abstract: Causal tracing of factual recall has been studied predominantly in dense transformer language models, where interventions localize information flow to l

local-aiarxiv-cs-cl
3 Jun 2026
← Previous
1…134135136137138…1262
Next →