AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,536 results
Safety

Breaking the Tokenizer Barrier: On-Policy Distillation across Model Families

DGX agent

arXiv:2606.09456v1 Announce Type: new Abstract: On-Policy Distillation (OPD) has become a core technique in the post-training of Large Language Models (LLMs) for transferring knowledge from domain exp

safetyarxiv-cs-lg
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Bridging Traditional Explainability Methods and Multimodal Multilingual Models: An XAI-Based Analysis

DGX agent

arXiv:2606.07533v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) effectively integrate text and audio to interpret context in complex interactive dialogues. However, the inte

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Coarse-to-Fine Hierarchical Alignment for UAV-based Human Detection using Diffusion Models

DGX agent

arXiv:2512.13869v3 Announce Type: replace Abstract: Training object detectors demands extensive, task-specific annotations, yet this requirement becomes impractical in UAV-based human detection due to

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models

DGX agent

arXiv:2606.09142v1 Announce Type: cross Abstract: Egocentric vision offers a first-person view of human perception and decision making, yet its potential for traffic-safety prediction remains underexp

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Diverse Thinking Schemata Elicit Better Reasoning in Large Language Models

DGX agent

arXiv:2606.08974v1 Announce Type: new Abstract: Large reasoning models (LRMs) have attracted increasing attention for their ability to solve complex mathematical problems by generating extended reason

safetyarxiv-cs-ai
9 Jun 2026
Research

Do Video Foundation Models Understand Intuitive Physics? A Layerwise Probing Analysis

DGX agent

arXiv:2606.09646v1 Announce Type: cross Abstract: We study whether pretrained video foundation models encode intuitive-physics information in their frozen representations, and how this information var

researcharxiv-cs-ai
9 Jun 2026
Research

DynaCF: Mitigating Shortcut Learning in Reward Models via Dynamic Counterfactual Sensitivity

DGX agent

arXiv:2606.09043v1 Announce Type: new Abstract: Reward models trained from pairwise preferences often exploit superficial shortcut cues rather than learning true response quality. We propose DynaCF, a

researcharxiv-cs-lg
9 Jun 2026
Research

Echo-Memory: A Controlled Study of Memory in Action World Models

DGX agent

arXiv:2606.09803v1 Announce Type: new Abstract: We present extbf{Echo-Memory}, a controlled study of memory mechanisms in action-conditioned world models. These models generate multi-segment videos fr

researcharxiv-cs-cv
9 Jun 2026
Tutorials

Emergence of Context Characteristics Sensitivity in Large Language Models

DGX agent

arXiv:2606.09525v1 Announce Type: cross Abstract: During instruction fine-tuning (IFT), large language models (LLMs) learn to follow instructions by using the provided context to answer a query. While

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

Enhancing Spatial Reasoning in Large Language Models for Metal-Organic Frameworks Structure Prediction

DGX agent

arXiv:2601.09285v2 Announce Type: replace Abstract: Metal-organic frameworks (MOFs) are porous crystalline materials with broad applications such as carbon capture and drug delivery, yet accurately pr

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Fable is a step-change in models, and I hope it changes how you work with Claude. More to come in a series of posts on how it’s reshaped our…

DGX agent

Fable is a step-change in models, and I hope it changes how you work with Claude. More to come in a series of posts on how it’s reshaped our work, but the TLDR: it’s time to be more ambitious. Claude

model-releasesthariq--x
9 Jun 2026
Local Ai

Frame Adjustments Demo: Most people think AI filmmaking means finding the one perfect model. It doesn't. @heydoughogan breaks down why the b…

DGX agent

Frame Adjustments Demo: Most people think AI filmmaking means finding the one perfect model. It doesn't. @heydoughogan breaks down why the best workflows are mix-and-match. Different models for differ

local-aicomfyui--x
9 Jun 2026
Research

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling

DGX agent

arXiv:2606.08302v1 Announce Type: new Abstract: Visual Autoregressive (VAR) models adopt a next-scale prediction paradigm, offering high-quality generation with substantially fewer decoding steps. How

researcharxiv-cs-cv
9 Jun 2026
Industry

HF has been an amazing partner since day one, so this was easy. As we’ve grown from a post-training shop into a full model lab, @huggingface…

DGX agent

HF has been an amazing partner since day one, so this was easy. As we’ve grown from a post-training shop into a full model lab, @huggingface is the obvious partner to scale the infra open models deman

industryclem-delangue--x
9 Jun 2026
Model Releases

How Much Dense Attention is Necessary? Oracle-Guided Sparse Prefill for Full/GQA Layers in Hybrid Long-Context Models

DGX agent

arXiv:2606.07703v1 Announce Type: cross Abstract: Long-context prefill remains expensive because full/GQA layers still score the historical sequence, even in hybrid models with local, sparse, linear,

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

How Well Do Latent World Models Understand Partially Observable Safety Constraints?

DGX agent

arXiv:2510.06492v2 Announce Type: replace Abstract: Latent world models are a promising approach for learning state representations and dynamics directly from high-dimensional observations, enabling r

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

IDEQ -- Improving Diffusion Models for the Traveling Salesman Problem (TSP) by Leveraging the Structure of the Solution Space

DGX agent

arXiv:2412.13858v2 Announce Type: replace Abstract: We investigate diffusion models to solve the Traveling Salesman Problem. Building on the recent DIFUSCO and T2TCO approaches, we propose IDEQ. IDEQ

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

IMUG-Bench: Benchmarking Unified Multimodal Models on Interleaved Understanding and Generation

DGX agent

arXiv:2606.09169v1 Announce Type: new Abstract: In recent years, unified multimodal models (UMMs) have emerged to support both understanding and generation within a single framework. Mastering dynamic

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Introducing Cohere's first open-source coding model: North Mini Code Small & efficient, designed for agentic performance and built for commu…

DGX agent

Cohere released North Mini Code, an open-source coding model designed to be small and efficient while optimizing for agentic performance and community use. The model represents Cohere's initial offeri

agentsclem-delangue--x
9 Jun 2026
Model Releases

Knowledge-Inclusive Adaptive Physics-Informed Neural Network for Microbial Interaction Modelling

DGX agent

arXiv:2606.07686v1 Announce Type: cross Abstract: Physics-Informed Neural Network (PINN) is a way of including knowledge in the form of equations in Machine Learning methods. Beyond equations, knowled

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

LargeMonitor: Monitoring Online Task-Free Continual Learning via Large Pretrained Models

DGX agent

arXiv:2606.09430v1 Announce Type: cross Abstract: Online task-free continual learning (TFCL) requires intelligent agents to sequentially accumulate knowledge from an unbounded, non-stationary data str

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Late-Layer Fusion is Enough: Dual-Path Vision Token Routing for Multimodal Large Language Models under Visual Saturation

DGX agent

arXiv:2606.09131v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) commonly inherit the deep, symmetric Transformer backbone designed for unimodal text modeling, and apply the sa

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Model-Based Learning of Whittle indices

DGX agent

arXiv:2511.20397v2 Announce Type: replace Abstract: We present BLINQ, a new model-based algorithm that learns the Whittle indices of an indexable, communicating and unichain Markov Decision Process (M

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Pretrained, Frozen, Still Leaking: Auditing Cross-Encoder Attribute Transfer in EEG Foundation Models

DGX agent

arXiv:2606.09189v1 Announce Type: cross Abstract: EEG foundation-model releases are usually audited one endpoint at a time: raw-reconstruction, membership inference, identity linkage, or DP-SGD on the

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Prisma-World: Camera-Controllable Multi-Agent Video World Model

DGX agent

arXiv:2606.09507v1 Announce Type: new Abstract: Video world models have made rapid progress in generating controllable visual experiences, but most of them still simulate the world from a single obser

safetyarxiv-cs-cv
9 Jun 2026
Safety

PROBE-Web: An Interactive System for Probing Evaluation Landscapes of Knowledge Graph Completion Models

DGX agent

arXiv:2606.08926v1 Announce Type: new Abstract: Knowledge graph completion (KGC) models are commonly evaluated using rank-based metrics such as MRR and Hits@K, despite different users often requiring

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

ProbeAct: Probe-Guided Training-Free Failure Recovery in Vision-Language-Action Models

DGX agent

arXiv:2606.09740v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models demonstrate strong perfor-1 mance on language-conditioned robotic manipulation within their training dis-2 tribution

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

Setting a custom price for a model in AgentsView

DGX agent

TIL: Setting a custom price for a model in AgentsView I've been really enjoying AgentsView by Wes McKinney as a tool for exploring my token usage across different coding agents running on my laptop. C

model-releasessimon-willison
9 Jun 2026
Research

SlideCheck: Guiding Self-Supervised Pretraining of Pathology Foundation Models via Dataset Distributions

DGX agent

arXiv:2606.07590v1 Announce Type: cross Abstract: Pathology foundation models are pretrained on large streams of WSI-derived patches, while supervision during data construction is often slide-level, s

researcharxiv-cs-ai
9 Jun 2026
Safety

Sycophancy as a Multilingual Alignment Failure: How Safety Degrades Across Languages, Topics, and Models

DGX agent

arXiv:2606.08451v1 Announce Type: cross Abstract: Safety-aligned large language models often exhibit sycophancy, which is the tendency to affirm users' opinions regardless of factual accuracy. Althoug

safetyarxiv-cs-ai
9 Jun 2026
Safety

Targeting World Models to Compromise Robot Learning Pipelines

DGX agent

arXiv:2606.09499v1 Announce Type: cross Abstract: World models have recently seen a rapid growth in both their popularity and capability as more data efficient tools for generating robot training data

safetyarxiv-cs-ai
9 Jun 2026
Safety

The ACUTE Protocol: Operationalizing Language Model Activations for Better Calibration, Utility, and Trust

DGX agent

arXiv:2606.07822v1 Announce Type: cross Abstract: As language models improve and become increasingly deployed to solve a variety of tasks, trustworthiness becomes essential. Calibration is a good prox

safetyarxiv-cs-ai
9 Jun 2026
Research

Transition-Based Digital Twin Modelling for Alzheimer's Disease under Sparse Longitudinal Data

DGX agent

arXiv:2606.09671v1 Announce Type: cross Abstract: Alzheimer's disease (AD) progression is highly heterogeneous and is typically observed through sparse and irregular longitudinal data, posing challeng

researcharxiv-cs-ai
9 Jun 2026
Safety

Video Understanding by Design: How Datasets Shape Video Models

DGX agent

arXiv:2509.09151v2 Announce Type: replace-cross Abstract: Research in video understanding has advanced rapidly, driven by increasingly diverse datasets and more powerful model architectures. While exi

safetyarxiv-cs-ai
9 Jun 2026
Safety

AdaJudge: Adaptive Multi-Perspective Judging for Reward Modeling

DGX agent

arXiv:2601.08097v2 Announce Type: replace Abstract: Reward modeling is essential for aligning large language models with human preferences, yet predominant architectures rely on a static pooling strat

safetyarxiv-cs-cl
8 Jun 2026
Safety

Agentic World Modeling for 6G: Near-Real-Time Generative State-Space Reasoning

DGX agent

arXiv:2511.02748v2 Announce Type: replace-cross Abstract: We argue that sixth-generation (6G) intelligence is not fluent token prediction but the capacity to imagine and choose -- to simulate future s

safetyarxiv-cs-lg
8 Jun 2026
Model Releases

Benchmarking Language Modeling for Lossless Compression of Full-Fidelity Audio

DGX agent

arXiv:2603.08683v2 Announce Type: replace-cross Abstract: Autoregressive 'language' models (LMs) trained on raw waveforms can be repurposed for lossless audio compression, but prior work is limited to

model-releasesarxiv-cs-ai
8 Jun 2026
Research

CountsDiff: A Diffusion Model on the Natural Numbers for Generation and Imputation of Count-Based Data

DGX agent

arXiv:2604.03779v2 Announce Type: replace-cross Abstract: Diffusion models have excelled at generative tasks for both continuous and token-based domains, but their application to discrete ordinal data

researcharxiv-cs-ai
8 Jun 2026
Agents

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a…

DGX agent

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a much broader set of verticalized agents and infrastructure.

agentsjerry-liu--x
8 Jun 2026
Model Releases

Inheritance Between Feedforward and Convolutional Networks via Model Projection

DGX agent

arXiv:2602.06245v2 Announce Type: replace-cross Abstract: Neural-network techniques are often transferred across architecture families by analogy, but such transfer is valid only when the assumptions

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Interesting model here 35b a3b trained for agentic use It gets 60.7 on Terminal Bench2 qwen 3.6 27b gets 59.3 Essentially the same Going to …

DGX agent

Interesting model here 35b a3b trained for agentic use It gets 60.7 on Terminal Bench2 qwen 3.6 27b gets 59.3 Essentially the same Going to have to try it out https://huggingface.co/nex-agi/Nex-N2-min

model-releasesclem-delangue--x
8 Jun 2026
Safety

LARA: Latent Action Representation Alignment for Vision-Language-Action Models

DGX agent

arXiv:2606.07100v1 Announce Type: new Abstract: Visual-language action (VLA) models enable robots to predict actions directly from observations and language instructions, but their performance depends

safetyarxiv-cs-cv
8 Jun 2026
Safety

Latent-space Attacks for Refusal Evasion in Language Models

DGX agent

arXiv:2605.21706v2 Announce Type: replace Abstract: Safety-aligned language models are trained to refuse harmful requests, yet refusal behavior can be suppressed by steering their internal representat

safetyarxiv-cs-ai
8 Jun 2026
Industry

Model routing is growing a lot these days

DGX agent

Model routing is growing a lot these days Good take My guess is - demand for intelligence is near infinite - but 80% of workloads will be running on 99% cheaper models within 12-18 months - 20% of wor

industryclem-delangue--x
8 Jun 2026
Model Releases

Seeing a number of benchmarks showing Opus is the best model for long-running work. Five tips for running Opus autonomously for hours/days: …

DGX agent

Seeing a number of benchmarks showing Opus is the best model for long-running work. Five tips for running Opus autonomously for hours/days: 1. Use auto mode for permissions, so Claude doesn’t ask for

model-releasesboris-cherny--x
8 Jun 2026
Applications

SpectCount: Spectrotemporal Counting via Synthetic Signals Improves Large Audio Language Models

DGX agent

arXiv:2606.06907v1 Announce Type: cross Abstract: Large audio language models (LALMs) extend large language models with an audio encoder and large-scale audio data. However, the scarcity of high-quali

applicationsarxiv-cs-ai
8 Jun 2026
Safety

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models

DGX agent

arXiv:2602.02600v3 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have recently emerged as a competitive alternative to autoregressive (AR) models, offering parallel decoding,

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

Textual Supervision Enhances Geospatial Representations in Vision-Language Models

DGX agent

arXiv:2606.07172v1 Announce Type: cross Abstract: Geospatial understanding is a critical yet underexplored dimension in the development of machine learning systems for tasks such as image geolocation

model-releasesarxiv-cs-ai
8 Jun 2026
← Previous
1…133134135136137…1262
Next →