AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlog
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,082 results
Local Ai

Batching for vision models is now available in Beta with our latest MLX engine update 👾 The updated engine also brings major improvements t…

DGX agent

Batching for vision models is now available in Beta with our latest MLX engine update 👾 The updated engine also brings major improvements to caching for faster inference overall. Turn on Developer Mod

local-ailm-studio--x
14 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

CLIP Tricks You: Training-free Token Pruning for Efficient Pixel Grounding in Large VIsion-Language Models

DGX agent

arXiv:2605.13178v1 Announce Type: cross Abstract: In large vision-language models, visual tokens typically constitute the majority of input tokens, leading to substantial computational overhead. To ad

researcharxiv-cs-ai
14 May 2026
Safety

Constraints-of-Thought: A Framework for Constrained Reasoning in Language-Model-Guided Search

DGX agent

arXiv:2510.08992v3 Announce Type: replace Abstract: While researchers have made significant progress in enabling large language models (LLMs) to perform multi-step planning, LLMs struggle to ensure th

safetyarxiv-cs-lg
14 May 2026
Industry

Dear AI labs, A neverending black and white kanban board modeled after Jira (no offense) is not what we want as the future of work. Give me …

DGX agent

Dear AI labs, A neverending black and white kanban board modeled after Jira (no offense) is not what we want as the future of work. Give me flexibility. Give me delight. Give me color. Please create t

industryallie-k--miller--x
14 May 2026
Safety

DisaBench: A Participatory Evaluation Framework for Disability Harms in Language Models

DGX agent

arXiv:2605.12702v1 Announce Type: new Abstract: General-purpose safety benchmarks for large language models do not adequately evaluate disability-related harms. We introduce DisaBench: a taxonomy of t

safetyarxiv-cs-ai
14 May 2026
Research

Dual-Pathway Circuits of Object Hallucination in Vision-Language Models

DGX agent

arXiv:2605.13156v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in bridging visual perception and natural language understanding, enabling a wid

researcharxiv-cs-cv
14 May 2026
Research

Generative Modeling by Minimizing the Wasserstein-2 Loss

DGX agent

arXiv:2406.13619v4 Announce Type: replace-cross Abstract: This paper develops a generative model by minimizing the second-order Wasserstein loss (the W_2 loss) through a distribution-dependent ordinar

researcharxiv-cs-lg
14 May 2026
Local Ai

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models

DGX agent

arXiv:2605.13375v1 Announce Type: cross Abstract: In Vision-Language Models (VLMs), processing a massive number of visual tokens incurs prohibitive computational overhead. While recent training-aware

local-aiarxiv-cs-ai
14 May 2026
Local Ai

Learning to See What You Need: Gaze Attention for Multimodal Large Language Models

DGX agent

arXiv:2605.13080v1 Announce Type: new Abstract: When humans describe a visual scene, they do not process the entire image uniformly; instead, they selectively fixate on regions relevant to their inten

local-aiarxiv-cs-cv
14 May 2026
Applications

MILM: Large Language Models for Multimodal Irregular Time Series with Informative Sampling

DGX agent

arXiv:2605.13711v1 Announce Type: new Abstract: Multimodal irregular time series (MITS) consist of asynchronous and irregularly sampled observations from heterogeneous numerical and textual channels.

applicationsarxiv-cs-lg
14 May 2026
Agents

Model. Harness. Context. The 3 main components of agents. As you build more agents, context increasingly lives AGENTS.md, skills, policies, …

DGX agent

Model. Harness. Context. The 3 main components of agents. As you build more agents, context increasingly lives AGENTS.md, skills, policies, examples, + generated research files. Context needs its own

agentsharrison-chase--x
14 May 2026
Tutorials

Most teams can pick frontier models. Fewer can run them at production scale without hitting constraints in latency, throughput, and governan…

DGX agent

Most teams can pick frontier models. Fewer can run them at production scale without hitting constraints in latency, throughput, and governance. Fireworks AI on @Azure AI Foundry provides the inference

tutorialsfireworks-ai--x
14 May 2026
Research

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy

DGX agent

arXiv:2304.11193v2 Announce Type: replace-cross Abstract: Predicting the outcomes of robotic actions, often referred to as learning a world model, in complex environments remains a fundamental challen

researcharxiv-cs-ai
14 May 2026
Research

Multitask Multimodal Fusion with Tabular Foundation Models for Peak and Durability Prediction of Pertussis Booster Response

DGX agent

arXiv:2605.12852v1 Announce Type: new Abstract: Pertussis booster vaccination produces immune responses that vary widely across individuals in both peak magnitude and long-term durability. These two p

researcharxiv-cs-lg
14 May 2026
Research

On the Limits of Latent Reuse in Diffusion Models

DGX agent

arXiv:2605.13448v1 Announce Type: cross Abstract: Diffusion models are often trained in low-dimensional latent spaces, which are then reused for related but shifted datasets. In this work, we study wh

researcharxiv-cs-lg
14 May 2026
Safety

Pretraining Language Models with Subword Regularization: An Empirical Study of BPE Dropout in Low-Resource NLP

DGX agent

arXiv:2605.13436v1 Announce Type: cross Abstract: Subword regularization methods such as BPE dropout are typically applied only during fine-tuning, while pretraining is usually done with deterministic

safetyarxiv-cs-lg
14 May 2026
Research

SceneGraphVLM: Dynamic Scene Graph Generation from Video with Vision-Language Models

DGX agent

arXiv:2605.13667v1 Announce Type: new Abstract: Scene graph generation provides a compact structured representation for visual perception, but accurate and fast graph prediction from images and videos

researcharxiv-cs-cv
14 May 2026
Research

TimelineReasoner: Advancing Timeline Summarization with Large Reasoning Models

DGX agent

arXiv:2605.12518v1 Announce Type: cross Abstract: The proliferation of online news poses a challenge to extracting structured timelines from unstructured content. While recent studies have shown that

researcharxiv-cs-ai
14 May 2026
Industry

Toto 2.0 is here: Datadog AI's 5 open-weights forecasting models (4m-2.5B params) finally make scaling work for time series forecasting! #1 …

DGX agent

Toto 2.0 is here: Datadog AI's 5 open-weights forecasting models (4m-2.5B params) finally make scaling work for time series forecasting! #1 on BOOM, GIFT-Eval, and TIME. Weights/code Apache 2.0. 🔗 Rea

industryclem-delangue--x
14 May 2026
Tools

Try Rime Mist v3 in voice finder directly: https://findtherightvoice.com/rime-labs--rime-mist-v3 Model pages: http://www.together.ai/models/…

DGX agent

Try Rime Mist v3 in voice finder directly: https://findtherightvoice.com/rime-labs--rime-mist-v3 Model pages: http://www.together.ai/models/rime-mist-v3 http://www.together.ai/models/rime-mist-v3-omni

toolstogether-ai--x
14 May 2026
Safety

What to Ignore, What to React: Visually Robust RL Fine-Tuning of VLA Models

DGX agent

arXiv:2605.13105v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning has shown promise for Vision-Language-Action (VLA) models in robotic manipulation, but deployment-time visual sh

safetyarxiv-cs-ro
14 May 2026
Research

When to Think Fast and Slow? AMOR: Adaptive Entropy Gate for Hybrid Models

DGX agent

arXiv:2602.13215v2 Announce Type: replace Abstract: Recurrent-attention hybrids aim to combine the efficiency of recurrence with the expressivity of attention, but existing approaches typically apply

researcharxiv-cs-ai
14 May 2026
Agents

You can now power your Hermes Agent, if using OpenAI models, with codex as the runtime for the core tools that it offers, with the flip of a…

DGX agent

Nous Research announced that Hermes Agents powered by OpenAI models can now use Codex as the runtime for executing core tools, enabling improved tool execution capabilities with a simple configuration

agentsnous-research--x
14 May 2026
Safety

A Survey of On-Policy Distillation for Large Language Models

DGX agent

arXiv:2604.00626v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable stud

safetyarxiv-cs-cl
13 May 2026
Safety

Adaption, co-founded by ex-Cohere VP of AI research Sara Hooker, unveils AutoScientist, which can automate the research loop behind model training and alignment (Russell Brandom/TechCrunch)

DGX agent

Russell Brandom / TechCrunch: Adaption, co-founded by ex-Cohere VP of AI research Sara Hooker, unveils AutoScientist, which can automate the research loop behind model training and alignment — For yea

safetytechmeme
13 May 2026
Industry

China's move to block Meta's Manus acquisition challenges the 'Singapore washing' model, as major Chinese tech companies build significant Singapore presences (Owen Walker/Financial Times)

DGX agent

Owen Walker / Financial Times: China's move to block Meta's Manus acquisition challenges the “Singapore washing” model, as major Chinese tech companies build significant Singapore presences — Beijing'

industrytechmeme
13 May 2026
Safety

Cluster-Aware Neural Collapse Prompt Tuning for Long-Tailed Generalization of Vision-Language Models

DGX agent

arXiv:2605.11939v1 Announce Type: new Abstract: Prompt learning has emerged as an efficient alternative to fine-tuning pre-trained vision-language models (VLMs). Despite its promise, current methods s

safetyarxiv-cs-cv
13 May 2026
Applications

Config, which is building a data layer for robotics foundation models, raised a 27M seed at a 200M+ valuation led by Samsung Venture Investment (Kate Park/TechCrunch)

DGX agent

Kate Park / TechCrunch: Config, which is building a data layer for robotics foundation models, raised a 27M seed at a 200M+ valuation led by Samsung Venture Investment — Asia's push into physical AI i

applicationstechmeme
13 May 2026
Safety

Diffusion-State Policy Optimization for Masked Diffusion Language Models

DGX agent

arXiv:2602.06462v3 Announce Type: replace Abstract: Masked diffusion language models generate text through iterative masked-token filling, but terminal-only rewards on final completions provide coarse

safetyarxiv-cs-cl
13 May 2026
Research

Do Language Models Encode Knowledge of Linguistic Constraint Violations?

DGX agent

arXiv:2605.12055v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong linguistic performance, yet their internal mechanisms for producing these predictions remain unclear. We inv

researcharxiv-cs-cl
13 May 2026
Research

DP-{lambda}CGD: Efficient Noise Correlation for Differentially Private Model Training

DGX agent

arXiv:2601.22334v2 Announce Type: replace Abstract: Differentially private stochastic gradient descent (DP-SGD) is the gold standard for training machine learning models with formal differential priva

researcharxiv-cs-lg
13 May 2026
Applications

Efficient LLM-based Advertising via Model Compression and Parallel Verification

DGX agent

arXiv:2605.11582v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable potential in advertising scenarios such as ad creative generation and targeted advertising. However,

applicationsarxiv-cs-cl
13 May 2026
Safety

Enhancing Target-Guided Proactive Dialogue Systems via Conversational Scenario Modeling and Intent-Keyword Bridging

DGX agent

arXiv:2605.11964v1 Announce Type: new Abstract: A target-guided proactive dialogue system aims to steer conversations proactively toward pre-defined targets, such as designated keywords or specific to

safetyarxiv-cs-cl
13 May 2026
Tools

External content is scanned in parallel by ML classifiers and the BrowseSafe model before agents act on it. File connector data is encrypted…

DGX agent

External content is scanned in parallel by ML classifiers and the BrowseSafe model before agents act on it. File connector data is encrypted in transit and at rest, uploaded files automatically delete

toolsperplexity--x
13 May 2026
Applications

From Token to Token Pair: Efficient Prompt Compression for Large Language Models in Clinical Prediction

DGX agent

arXiv:2605.11774v1 Announce Type: new Abstract: By processing electronic health records (EHRs) as natural language sequences, large language models (LLMs) have shown potential in clinical prediction t

applicationsarxiv-cs-cl
13 May 2026
Industry

Google says Chromebooks will get support through their 'existing date commitment', and 'many' models are 'eligible to transition' to the Googlebook experience (Ben Schoon/9to5Google)

DGX agent

Ben Schoon / 9to5Google: Google says Chromebooks will get support through their “existing date commitment”, and “many” models are “eligible to transition” to the Googlebook experience — During The And

industrytechmeme
13 May 2026
Safety

Gradient-Free Noise Optimization for Reward Alignment in Generative Models

DGX agent

arXiv:2605.11347v1 Announce Type: cross Abstract: Existing reward alignment methods for diffusion and flow models rely on multi-step stochastic trajectories, making them difficult to extend to determi

safetyarxiv-cs-cv
13 May 2026
Applications

I don't understand the path forward for Mythos releases. Google & OpenAI will have equivalent models, and they are approaching AI cyber risk…

DGX agent

I don't understand the path forward for Mythos releases. Google & OpenAI will have equivalent models, and they are approaching AI cyber risk guardrails differently, so they will presumably just releas

applicationsethan-mollick--x
13 May 2026
Agents

Joint Learning of Hierarchical Neural Options and Abstract World Model

DGX agent

arXiv:2602.02799v2 Announce Type: replace Abstract: Building agents that can perform new skills by composing existing skills is a long-standing goal of AI agent research. Towards this end, we investig

agentsarxiv-cs-lg
13 May 2026
Safety

Model-based Bootstrap of Controlled Markov Chains

DGX agent

arXiv:2605.12410v1 Announce Type: cross Abstract: We propose and analyze a model-based bootstrap for transition kernels in finite controlled Markov chains (CMCs) with possibly nonstationary or history

safetyarxiv-cs-lg
13 May 2026
Applications

Model-Level GNN Explanations via Rule-to-Graph Readout for Logit Reconstruction

DGX agent

arXiv:2503.09051v2 Announce Type: replace Abstract: We propose a novel model-level GNN explanation framework that shifts the explanation target from class-wise rule extraction to rule-based logit reco

applicationsarxiv-cs-lg
13 May 2026
Industry

New paper: research agenda for secret loyalties Imagine a frontier model that has been trained to covertly advance a specific actor's intere…

DGX agent

New paper: research agenda for secret loyalties Imagine a frontier model that has been trained to covertly advance a specific actor's interests (a nation-state, a CEO, an adversary). @joemkwon argues

industryemad-mostaque--x
13 May 2026
Research

OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models

DGX agent

arXiv:2605.11803v1 Announce Type: new Abstract: As Video Large Language Models (Video-LLMs) scale to longer and more complex videos, their inference cost grows rapidly due to the large volume of visua

researcharxiv-cs-cv
13 May 2026
Applications

PayPal runs 74,000 weekly tasks in Perplexity Enterprise. Teams use it for model validation, channel performance, market trend research, com…

DGX agent

PayPal runs 74,000 weekly tasks in Perplexity Enterprise. Teams use it for model validation, channel performance, market trend research, competitive intelligence, and product analysis. Read the custom

applicationsperplexity--x
13 May 2026
Local Ai

Rank Is Not Capacity: Spectral Occupancy for Latent Graph Models

DGX agent

arXiv:2605.11142v1 Announce Type: new Abstract: Graph representation learning has become a standard approach for analyzing networked data, with latent embeddings widely used for link prediction, commu

local-aiarxiv-cs-lg
13 May 2026
Safety

Red-Teaming Text-to-Image Models via In-Context Experience Replay and Semantic-Preserving Prompt Rewriting

DGX agent

arXiv:2411.16769v3 Announce Type: replace-cross Abstract: Understanding the capabilities of text-to-image (T2I) models in harmful content generation is essential to safety and compliance. However, hum

safetyarxiv-cs-cl
13 May 2026
Safety

Simpson's Paradox in Behavioral Curves: How Aggregation Distorts Parametric Models of User Dynamics

DGX agent

arXiv:2605.11017v1 Announce Type: new Abstract: Behavioral curve modeling -- fitting parametric functions to engagement-versus-exposure data -- is standard practice in recommendation, advertising, and

safetyarxiv-cs-lg
13 May 2026
Safety

The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested

DGX agent

arXiv:2605.11496v1 Announce Type: cross Abstract: Recent published evidence from frontier laboratories shows that contemporary AI models can recognise evaluation contexts, latently represent them, and

safetyarxiv-cs-lg
13 May 2026
← Previous
1…277278279280281…1294
Next →