AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,036 results
Research

When Discourse Pressures Conflict: Information Structure in Vision-Language Model Outputs

DGX agent

arXiv:2605.28346v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly evaluated for whether they identify the right visual content, but little is known about whether they expr

researcharxiv-cs-cl
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

When pre-training hurts LoRA fine-tuning: a dynamical analysis via single-index models

DGX agent

arXiv:2602.02855v2 Announce Type: replace Abstract: Pre-training on a source task is usually expected to facilitate fine-tuning on similar downstream problems. In this work, we mathematically show tha

safetyarxiv-cs-lg
28 May 2026
Tutorials

Which Pretraining Paradigm Better Serves Spatial Intelligence? An Empirical Comparison of Vision-Language and Video Generation Models

DGX agent

arXiv:2605.28132v1 Announce Type: new Abstract: Spatial intelligence requires visual representations that capture both semantic objects and geometric structure in the physical world. To support this,

tutorialsarxiv-cs-cv
28 May 2026
Research

Would you like to join the research effort on JEPA and World Models easily? After a full year of hard work, we’re excited to finally release…

DGX agent

Would you like to join the research effort on JEPA and World Models easily? After a full year of hard work, we’re excited to finally release stable-worldmodel: an open-source, scalable platform built

researchyann-lecun--x
28 May 2026
Safety

Beyond Pairwise Preferences: Listwise Reward-Aware Alignment for Diffusion Models

DGX agent

arXiv:2605.26491v1 Announce Type: cross Abstract: Preference optimization has emerged as an efficient alternative to online reinforcement learning from human feedback (RLHF) for aligning text-to-image

safetyarxiv-cs-cv
27 May 2026
Research

Chartographer: Counterfactual Chart Generation for Evaluating Vision-Language Models

DGX agent

arXiv:2605.27311v1 Announce Type: new Abstract: Chart question-answering (QA) benchmarks aim to pose questions that require visual reasoning to correctly answer, but models can often reach solutions t

researcharxiv-cs-cl
27 May 2026
Model Releases

CodecCap: High-Fidelity Codec-Inspired Residual Modeling for Dense Video Captioning

DGX agent

arXiv:2605.26967v1 Announce Type: new Abstract: Existing video captioning methods struggle to balance visual fidelity and redundancy: holistic captions are compact but lose fine-grained evidence, wher

model-releasesarxiv-cs-cv
27 May 2026
Research

ContextGuard: Structured Self-Auditing for Context Learning in Language Models

DGX agent

arXiv:2605.26827v1 Announce Type: cross Abstract: Recent benchmarks reveal that despite strong reasoning capabilities, large language models (LLMs) still struggle to faithfully apply complex contextua

researcharxiv-cs-ai
27 May 2026
Safety

Cultural Value Alignment Via Latent Activation Steering in Large Language Models

DGX agent

arXiv:2605.26365v1 Announce Type: new Abstract: Large Language Models (LLMs) often exhibit homogenized cultural perspectives. While the World Values Survey (WVS) provides a gold standard for mapping h

safetyarxiv-cs-cl
27 May 2026
Local Ai

Edge AI Deployment Beyond Models: A BSP-Aware Systems Framework for Industrial Embedded Platforms

DGX agent

arXiv:2605.26119v1 Announce Type: cross Abstract: Industrial Edge AI programs often begin with the model and only later confront the platform. That sequencing is attractive because it allows early dem

local-aiarxiv-cs-ai
27 May 2026
Research

Emergent Causal-Geometric Dynamics Across Depth in Large Language Models

DGX agent

arXiv:2602.04931v2 Announce Type: replace-cross Abstract: Geometric analyses of large language model (LLM) representations reveal structured variation across depth but remain fundamentally correlation

researcharxiv-cs-ai
27 May 2026
Safety

Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights

DGX agent

arXiv:2501.06708v5 Announce Type: replace-cross Abstract: Large-scale web-crawled datasets contain noise, bias, and irrelevant information, necessitating data selection techniques. Existing methods de

safetyarxiv-cs-ai
27 May 2026
Applications

Google has the only true Omni model, but the elements aren't hooked up. It appears it can take in & output audio, images. video, songs, text…

DGX agent

Google has the only true Omni model, but the elements aren't hooked up. It appears it can take in & output audio, images. video, songs, text, code, etc. But right now each type of output is separate.

applicationsethan-mollick--x
27 May 2026
Industry

Great to see @poolsideai (US lab) committing to open sourcing their foundation models going forward Laguna is an interesting release, check …

DGX agent

Great to see @poolsideai (US lab) committing to open sourcing their foundation models going forward Laguna is an interesting release, check it out @Shaughnessy119 https://poolside.ai/blog/introducing-

industryemad-mostaque--x
27 May 2026
Industry

If the Founder of Hugging Face asks, you gotta do it. Models and dataset now live: https://huggingface.co/papers/2605.22391 Also built an ex…

DGX agent

If the Founder of Hugging Face asks, you gotta do it. Models and dataset now live: https://huggingface.co/papers/2605.22391 Also built an explorer: https://huggingface.co/spaces/Kaikaku/epicure-explor

industryclem-delangue--x
27 May 2026
Tutorials

Learning Energy-Based Models from Stochastic Interpolants using Spatiotemporal Differences

DGX agent

arXiv:2605.26850v1 Announce Type: new Abstract: Learning an energy-based model from data samples is a central problem in machine learning. Many recent and popular methods, such as denoising score matc

tutorialsarxiv-cs-lg
27 May 2026
Tutorials

Learning to Diagnose and Correct Errors: Towards Moral Sensitivity Acquisition in Large Language Models

DGX agent

arXiv:2601.03079v4 Announce Type: replace Abstract: Moral sensitivity is the most fundamental capability underlying human moral competence. Although many approaches aim to align large language models

tutorialsarxiv-cs-cl
27 May 2026
Research

Left-Right Symmetry Breaking in CLIP-style Vision-Language Models Trained on Synthetic Spatial-Relation Data

DGX agent

arXiv:2601.12809v2 Announce Type: replace-cross Abstract: Spatial understanding remains a key challenge in vision-language models. Yet it is still unclear whether such understanding is truly acquired,

researcharxiv-cs-ai
27 May 2026
Research

LUCoS: Latent Unsupervised Context Selection for Tabular Foundation Models

DGX agent

arXiv:2605.27254v1 Announce Type: cross Abstract: Selecting which instances to label is a key challenge in low-label tabular learning. For recent Tabular Foundation Models such as TabPFN, context sele

researcharxiv-cs-ai
27 May 2026
Tutorials

Model Unlearning Objectives Vary for Distinct Language Functions

DGX agent

arXiv:2605.26454v1 Announce Type: new Abstract: Large language models (LLMs) learn undesirable properties during pretraining, including dangerous knowledge and toxic text generation. Just as post-trai

tutorialsarxiv-cs-cl
27 May 2026
Safety

Olaf-World: Orienting Latent Actions for Video World Modeling

DGX agent

arXiv:2602.10104v2 Announce Type: replace-cross Abstract: Scaling action-controllable world models is limited by the scarcity of action labels. While latent action learning promises to extract control

safetyarxiv-cs-ai
27 May 2026
Research

Respecting Modality Gap in Post-hoc Out-of-distribution Detection with Pre-trained Vision-Language Models

DGX agent

arXiv:2605.26661v1 Announce Type: cross Abstract: Out-of-distribution (OOD) detection has emerged as a popular technique to enhance the reliability of machine learning models by identifying unexpected

researcharxiv-cs-ai
27 May 2026
Model Releases

SeDT: Sentence-Transformer Decision-Transformer Conditioning for Multi-Turn Conversation Reliability

DGX agent

arXiv:2605.26788v1 Announce Type: cross Abstract: Large language models (LLMs) achieve impressive performance when a task is fully specified in a single turn, yet the same models lose up to 39% of tha

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning

DGX agent

arXiv:2603.28730v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown impressive capabilities across diverse tasks, motivating efforts to leverage these models to supervis

model-releasesarxiv-cs-cl
27 May 2026
Local Ai

The Constraint Tax: Measuring Validity-Correctness Tradeoffs in Structured Outputs for Small Language Models

DGX agent

arXiv:2605.26128v1 Announce Type: new Abstract: Production LLM systems increasingly require machine-readable outputs: JSON objects, typed traces, regex-constrained fields, and tool-call schemas. This

local-aiarxiv-cs-lg
27 May 2026
Safety

To model human linguistic prediction, make LLMs less superhuman

DGX agent

arXiv:2510.05141v2 Announce Type: replace Abstract: When we read, we make predictions about upcoming words; these predictions influence our reading behavior. The success of large language models (LLMs

safetyarxiv-cs-cl
27 May 2026
Tutorials

Wow. It looks like the @XiaomiMiMo v2.5 model is insanely good value :O (Price for each prompt shown after each answer. Context includes >40…

DGX agent

Jeremy Howard comments on the Xiaomi MiMo v2.5 model, highlighting its exceptional value proposition and cost-effectiveness for prompt processing. The post appears to include comparative pricing data

tutorialsjeremy-howard--x
27 May 2026
Safety

A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot Segmentation, Classification, and Deblurring

DGX agent

arXiv:2605.26026v1 Announce Type: cross Abstract: Light sheet fluorescence microscopy (LSM) enables high-resolution, three-dimensional (3D) imaging of biological specimens, providing rich volumetric d

safetyarxiv-cs-ai
26 May 2026
Agents

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery

DGX agent

arXiv:2604.05550v2 Announce Type: replace Abstract: Artificial intelligence research increasingly depends on prolonged cycles of reproduction, debugging, and iterative refinement to achieve State-Of-T

agentsarxiv-cs-cl
26 May 2026
Safety

Capability and Robustness Cannot Both Be Free: An Information-Theoretic Bound for Vision-Language-Action Models

DGX agent

arXiv:2605.25889v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly deployed on real robots, where each predicted action is executed and each failure carries a safet

safetyarxiv-cs-lg
26 May 2026
Local Ai

Comfy will be at @aionthelot Our CEO @yoland_yan is joining a panel of AI video leaders for a progress report on the state of video models -…

DGX agent

Comfy will be at @aionthelot Our CEO @yoland_yan is joining a panel of AI video leaders for a progress report on the state of video models - where they excel, where they fall short, and what's coming

local-aicomfyui--x
26 May 2026
Research

Confidence Calibration in Large Language Models

DGX agent

arXiv:2605.23909v1 Announce Type: new Abstract: We investigate the calibration of large language models' (LLMs') confidence across diverse tasks. The results of our preregistered study show that the c

researcharxiv-cs-ai
26 May 2026
Safety

Dynamic Optimization and Safety Indicator Injection for Jailbreaking Text-to-Image Models with Multimodal Safety Filters

DGX agent

arXiv:2505.18979v2 Announce Type: replace Abstract: Text-to-image (T2I) models can generate not-safe-for-work (NSFW) content, motivating multi-stage safety pipelines with both text and image filters.

safetyarxiv-cs-lg
26 May 2026
Safety

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving

DGX agent

arXiv:2512.04733v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving (AD) systems increasingly adopt vision-language-action (VLA) models, yet they typically ignore the passenger's e

safetyarxiv-cs-ai
26 May 2026
Model Releases

ERNIE-Image Technical Report

DGX agent

arXiv:2605.25347v1 Announce Type: cross Abstract: We introduce ERNIE-Image, an open-source text-to-image generation model built upon an 8B single-stream DiT architecture. ERNIE-Image aims to bridge th

model-releasesarxiv-cs-lg
26 May 2026
Tutorials

EXPO-FT: Sample-Efficient Reinforcement Learning Finetuning for Vision-Language-Action Models

DGX agent

arXiv:2605.25477v1 Announce Type: cross Abstract: The ability to efficiently and reliably learn new tasks has been a foundational challenge in robotics. Vision-Language-Action (VLA) models have demons

tutorialsarxiv-cs-ai
26 May 2026
Model Releases

Hadamard Representation: Scaffolding Performance Across Model-free RL

DGX agent

arXiv:2406.09079v5 Announce Type: replace Abstract: Deep reinforcement learning agents progressively lose representational capacity during training: neurons become dormant, removing active capacity fr

model-releasesarxiv-cs-lg
26 May 2026
Applications

Hypothesis Generation and Inductive Inference in Children and Language Models

DGX agent

arXiv:2605.24528v1 Announce Type: new Abstract: Real world decision-making requires constructing mental models under uncertainty over evidence, over the underlying causal rules, and over the state of

applicationsarxiv-cs-ai
26 May 2026
Safety

Jailbreak to Protect: Buffering and Reinforcing via Temporary Jailbreaking for Safe Fine-Tuning in Large Language Models

DGX agent

arXiv:2605.24550v1 Announce Type: new Abstract: Fine-tuning-as-a-Service (FaaS) enables personalization of large language models (LLMs), but it can weaken safety-alignment under harmful fine-tuning at

safetyarxiv-cs-ai
26 May 2026
Tutorials

// Language Models Need Sleep // Let your agents 'sleep', folks. On a serious note, this is a fascinating paper on getting the most from lon…

DGX agent

// Language Models Need Sleep // Let your agents 'sleep', folks. On a serious note, this is a fascinating paper on getting the most from long-horizon agents. Here is the problem with agents today: Att

tutorialsdair-ai--x
26 May 2026
Model Releases

Learning to Reason Efficiently with A* Post-Training

DGX agent

arXiv:2605.24597v1 Announce Type: new Abstract: Many applications of large language models (LLMs) require deductive reasoning, yet models frequently produce incorrect or redundant inference steps. We

model-releasesarxiv-cs-ai
26 May 2026
Research

Logic-Guided Vector Fields for Constrained Generative Modeling

DGX agent

arXiv:2602.02009v2 Announce Type: replace Abstract: Neuro-symbolic systems aim to combine the expressive structure of symbolic logic with the flexibility of neural learning; yet, generative models typ

researcharxiv-cs-lg
26 May 2026
Safety

MAGIC: Multimodal Alignment & Grounding-aware Instruction Coreset for Vision-Language Models

DGX agent

arXiv:2605.26004v1 Announce Type: cross Abstract: Instruction tuning of large vision-language models (LVLMs) increasingly depends on massive multimodal corpora, yet these datasets contain samples with

safetyarxiv-cs-cl
26 May 2026
Agents

MCPXKIT: The Unified Toolkit for Analyzing Model Context Protocol Security

DGX agent

arXiv:2508.12538v2 Announce Type: replace-cross Abstract: The Model Context Protocol (MCP) has emerged as a universal standard that enables AI agents to seamlessly connect with external tools, signifi

agentsarxiv-cs-ai
26 May 2026
Research

Mitigating Object Hallucinations in Vision-Language Models through Region-Aware Attention Recalibration

DGX agent

arXiv:2605.24957v1 Announce Type: new Abstract: The generation of factually incorrect objects, commonly known as object hallucination, remains a persistent challenge in Large Vision-Language Models (L

researcharxiv-cs-ai
26 May 2026
Industry

OpenRouter raised 113M led by CapitalG, a source says at a 1.3B valuation, and now processes 25T tokens across 400+ models weekly, up from 5T six months ago (Michael J. de la Merced/New York Times)

DGX agent

Michael J. de la Merced / New York Times: OpenRouter raised 113M led by CapitalG, a source says at a 1.3B valuation, and now processes 25T tokens across 400+ models weekly, up from 5T six months ago —

industrytechmeme
26 May 2026
Local Ai

Regional Condition Custom Node for Anima model

DGX agent

Anima is a 2 billion parameter text-to-image model focused mainly on anime concepts and styles, but also capable of generating other non-photorealistic content. A Regional Condition Custom Node for An

local-air-stablediffusion
26 May 2026
Safety

Side-by-side Comparison Amplifies Dialect Bias in Language Models

DGX agent

arXiv:2605.24384v1 Announce Type: cross Abstract: Language models (LMs) can exhibit systematic biases against speakers based on variations in their dialects, even in the absence of a dialect label, a

safetyarxiv-cs-ai
26 May 2026
← Previous
1…230231232233234…1272
Next →