AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,521 results
Applications

Adaptive Action Chunking at Inference-time for Vision-Language-Action Models

DGX agent

arXiv:2604.04161v2 Announce Type: replace Abstract: In Vision-Language-Action (VLA) models, action chunking (i.e., executing a sequence of actions without intermediate replanning) is a key technique t

applicationsarxiv-cs-ro
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

AssemLM: Spatial Reasoning Multimodal Large Language Models for Robotic Assembly

DGX agent

arXiv:2604.08983v1 Announce Type: new Abstract: Spatial reasoning is a fundamental capability for embodied intelligence, especially for fine-grained manipulation tasks such as robotic assembly. While

model-releasesarxiv-cs-ro
13 Apr 2026
Model Releases

AudioGuard: Toward Comprehensive Audio Safety Protection Across Diverse Threat Models

DGX agent

arXiv:2604.08867v1 Announce Type: cross Abstract: Audio has rapidly become a primary interface for foundation models, powering real-time voice assistants. Ensuring safety in audio systems is inherentl

model-releasesarxiv-cs-ai
13 Apr 2026
Research

BlendFusion -- Scalable Synthetic Data Generation for Diffusion Model Training

DGX agent

arXiv:2604.09022v1 Announce Type: new Abstract: With the rapid adoption of diffusion models, synthetic data generation has emerged as a promising approach for addressing the growing demand for large-s

researcharxiv-cs-cv
13 Apr 2026
Safety

Cards Against LLMs: Benchmarking Humor Alignment in Large Language Models

DGX agent

arXiv:2604.08757v1 Announce Type: cross Abstract: Humor is one of the most culturally embedded and socially significant dimensions of human communication, yet it remains largely unexplored as a dimens

safetyarxiv-cs-ai
13 Apr 2026
Tutorials

Decomposing the Delta: What Do Models Actually Learn from Preference Pairs?

DGX agent

arXiv:2604.08723v1 Announce Type: cross Abstract: Preference optimization methods such as DPO and KTO are widely used for aligning language models, yet little is understood about what properties of pr

tutorialsarxiv-cs-ai
13 Apr 2026
Safety

EvoLen: Evolution-Guided Tokenization for DNA Language Model

DGX agent

arXiv:2604.08698v1 Announce Type: new Abstract: Tokens serve as the basic units of representation in DNA language models (DNALMs), yet their design remains underexplored. Unlike natural language, DNA

safetyarxiv-cs-lg
13 Apr 2026
Research

From Navigation to Refinement: Revealing the Two-Stage Nature of Flow-based Diffusion Models through Oracle Velocity

DGX agent

arXiv:2512.02826v3 Announce Type: replace-cross Abstract: Flow-based diffusion models have emerged as a leading paradigm for training generative models across images and videos. However, their memoriz

researcharxiv-cs-ai
13 Apr 2026
Research

HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models

DGX agent

arXiv:2604.06165v2 Announce Type: replace Abstract: Large vision-language models can produce object hallucinations in image descriptions, highlighting the need for effective detection and mitigation s

researcharxiv-cs-cv
13 Apr 2026
Model Releases

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling

DGX agent

arXiv:2604.05072v2 Announce Type: replace Abstract: Recent large language models have shifted SVG generation from differentiable rendering optimization to autoregressive program synthesis. However, ex

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models

DGX agent

arXiv:2604.08970v1 Announce Type: cross Abstract: We study predictive multilingual evaluation: estimating how well a model will perform on a task in a target language when direct benchmark results are

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Multi-task Just Recognizable Difference for Video Coding for Machines: Database, Model, and Coding Application

DGX agent

arXiv:2604.09421v1 Announce Type: cross Abstract: Just Recognizable Difference (JRD) boosts coding efficiency for machine vision through visibility threshold modeling, but is currently limited to a si

researcharxiv-cs-cv
13 Apr 2026
Safety

Post-Hoc Guidance for Consistency Models by Joint Flow Distribution Learning

DGX agent

arXiv:2604.08828v1 Announce Type: cross Abstract: Classifier-free Guidance (CFG) lets practitioners trade-off fidelity against diversity in Diffusion Models (DMs). The practicality of CFG is however h

safetyarxiv-cs-cv
13 Apr 2026
Research

Predictive Entropy Links Calibration and Paraphrase Sensitivity in Medical Vision-Language Models

DGX agent

arXiv:2604.08941v1 Announce Type: new Abstract: Medical Vision Language Models VLMs suffer from two failure modes that threaten safe deployment mis calibrated confidence and sensitivity to question re

researcharxiv-cs-lg
13 Apr 2026
Local Ai

Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models

DGX agent

arXiv:2505.12509v3 Announce Type: replace-cross Abstract: Post-hoc explanations provide transparency and are essential for guiding model optimization, such as prompt engineering and data sanitation. H

local-aiarxiv-cs-ai
13 Apr 2026
Research

State Space Models are Effective Sign Language Learners: Exploiting Phonological Compositionality for Vocabulary-Scale Recognition

DGX agent

arXiv:2604.08761v1 Announce Type: new Abstract: Sign language recognition suffers from catastrophic scaling failure: models achieving high accuracy on small vocabularies collapse at realistic sizes. E

researcharxiv-cs-cv
13 Apr 2026
Model Releases

The AI Codebase Maturity Model: From Assisted Coding to Self-Sustaining Systems

DGX agent

arXiv:2604.09388v1 Announce Type: cross Abstract: AI coding tools are widely adopted, but most teams plateau at prompt-and-review without a framework for systematic progression. This paper presents th

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Toward World Models for Epidemiology

DGX agent

arXiv:2604.09519v1 Announce Type: new Abstract: World models have emerged as a unifying paradigm for learning latent dynamics, simulating counterfactual futures, and supporting planning under uncertai

safetyarxiv-cs-lg
13 Apr 2026
Tutorials

Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models

DGX agent

arXiv:2604.09227v1 Announce Type: cross Abstract: Image generative models have become indispensable tools to yield exquisite high-resolution (HR) images for everyone, ranging from general users to pro

tutorialsarxiv-cs-cv
13 Apr 2026
Local Ai

Tried Ollama Cloud, just realize only Kimi model accept images

DGX agent

A Reddit user exploring Ollama Cloud noted that, at the time of their post, only the Kimi model supported image (vision/multimodal) inputs among the available cloud models. Kimi K2.5 is a native multi

local-air-ollama
13 Apr 2026
Research

Uncertainty-Aware Transformers: Conformal Prediction for Language Models

DGX agent

arXiv:2604.08885v1 Announce Type: new Abstract: Transformers have had a profound impact on the field of artificial intelligence, especially on large language models and their variants. However, as was

researcharxiv-cs-lg
13 Apr 2026
Model Releases

We conducted cyber evaluations of Claude Mythos Preview and found that it is the first model to complete an AISI cyber range end-to-end. 🧵

DGX agent

Boris Cherny and colleagues conducted cybersecurity evaluations of Claude Mythos Preview, finding it to be the first AI model to successfully complete an AISI (AI Safety Institute) cyber range end-to-

model-releasesboris-cherny--x
13 Apr 2026
Tutorials

How to train loras in One Trainer for Z Image using Civitai models?

DGX agent

This r/StableDiffusion post likely discusses the community challenge of using OneTrainer — a locally-installed, GUI-based LoRA training tool — to train LoRAs specifically compatible with Z Image (Tong

tutorialsr-stablediffusion
12 Apr 2026
Local Ai

@hwchase17 I think harness/managed agents is a way for Anthropic to keep its Moat. As models get mature, the need for cloud LLMs might reduc…

DGX agent

@hwchase17 I think harness/managed agents is a way for Anthropic to keep its Moat. As models get mature, the need for cloud LLMs might reduce and local LLM models might increase (save cost) and thus t

local-aiharrison-chase--x
12 Apr 2026
Model Releases

KIV: 1M token context window on a RTX 4070 (12GB VRAM), no retraining, drop-in HuggingFace cache replacement - Works with any model that uses DynamicCache [P]

DGX agent

KIV is a project shared on r/MachineLearning presenting a drop-in replacement for HuggingFace's `DynamicCache` that enables up to 1 million token context windows on consumer hardware with only 12GB of

model-releasesr-machinelearning
12 Apr 2026
Local Ai

What's model should I run?

DGX agent

A Reddit discussion from the r/ollama community where a user seeks advice on which AI language model to run locally using Ollama. Responses likely include hardware-based recommendations (such as RAM a

local-air-ollama
11 Apr 2026
Model Releases

A Benchmark of Classical and Deep Learning Models for Agricultural Commodity Price Forecasting on A Novel Bangladeshi Market Price Dataset

DGX agent

arXiv:2604.06227v1 Announce Type: new Abstract: Accurate short-term forecasting of agricultural commodity prices is critical for food security planning and smallholder income stabilisation in developi

model-releasesarxiv-cs-lg
10 Apr 2026
Applications

A Comparative Study of Demonstration Selection for Practical Large Language Models-based Next POI Prediction

DGX agent

arXiv:2604.06207v1 Announce Type: cross Abstract: This paper investigates demonstration selection strategies for predicting a user's next point-of-interest (POI) using large language models (LLMs), ai

applicationsarxiv-cs-ai
10 Apr 2026
Model Releases

AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors

DGX agent

arXiv:2601.20524v2 Announce Type: replace Abstract: Zero-shot anomaly detection aims to detect and localise abnormal regions in the image without access to any in-domain training images. While recent

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models

DGX agent

arXiv:2509.23435v2 Announce Type: replace-cross Abstract: The creation of high-quality multimodal datasets remains fundamental for advancing role-playing capabilities in large language models (LLMs).

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Bootstrapping Sign Language Annotations with Sign Language Models

DGX agent

arXiv:2604.07606v1 Announce Type: new Abstract: AI-driven sign language interpretation is limited by a lack of high-quality annotated data. New datasets including ASL STEM Wiki and FLEURS-ASL contain

model-releasesarxiv-cs-cv
10 Apr 2026
Tools

ChatGPT voice mode is a weaker model

DGX agent

I think it's non-obvious to many people that the OpenAI voice mode runs on a much older, much weaker model - it feels like the AI that you can talk to should be the smartest AI but it really isn't. If

toolssimon-willison
10 Apr 2026
Research

Compact Example-Based Explanations for Language Models

DGX agent

arXiv:2601.03786v2 Announce Type: replace Abstract: Training data influence estimation methods quantify the contribution of training documents to a model's output, making them a promising source of in

researcharxiv-cs-cl
10 Apr 2026
Local Ai

ConfusionPrompt: Practical Private Inference for Online Large Language Models

DGX agent

arXiv:2401.00870v5 Announce Type: replace-cross Abstract: State-of-the-art large language models (LLMs) are typically deployed as online services, requiring users to transmit detailed prompts to cloud

local-aiarxiv-cs-ai
10 Apr 2026
Model Releases

Detecting HIV-Related Stigma in Clinical Narratives Using Large Language Models

DGX agent

arXiv:2604.07717v1 Announce Type: new Abstract: Human immunodeficiency virus (HIV)-related stigma is a critical psychosocial determinant of health for people living with HIV (PLWH), influencing mental

model-releasesarxiv-cs-cl
10 Apr 2026
Research

DiffVC: A Non-autoregressive Framework Based on Diffusion Model for Video Captioning

DGX agent

arXiv:2604.08084v1 Announce Type: new Abstract: Current video captioning methods usually use an encoder-decoder structure to generate text autoregressively. However, autoregressive methods have inhere

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models

DGX agent

arXiv:2604.08284v1 Announce Type: new Abstract: Large language models store not only isolated facts but also rules that support reasoning across symbolic expressions, natural language explanations, an

model-releasesarxiv-cs-cl
10 Apr 2026
Applications

Environmental, Social and Governance Sentiment Analysis on Slovene News: A Novel Dataset and Models

DGX agent

arXiv:2604.06826v1 Announce Type: cross Abstract: Environmental, Social, and Governance (ESG) considerations are increasingly integral to assessing corporate performance, reputation, and long-term sus

applicationsarxiv-cs-ai
10 Apr 2026
Agents

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexib…

DGX agent

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexibility, but to ensure we actually own the long-term memory th

agentsharrison-chase--x
10 Apr 2026
Applications

HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models

DGX agent

arXiv:2512.09928v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently enabled robotic manipulation by grounding visual and linguistic cues into actions. However, most V

applicationsarxiv-cs-ro
10 Apr 2026
Tools

If you ask ChatGPT voice mode for its knowledge cutoff date it tells you April 2024 - it's a GPT-4o era model

DGX agent

ChatGPT's voice mode, when directly queried about its knowledge cutoff date, reports April 2024, indicating it is powered by a GPT-4o era model rather than a more recent one. GPT-4o initially had ...

toolssimon-willison--x
10 Apr 2026
Research

In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads

DGX agent

arXiv:2604.06356v1 Announce Type: cross Abstract: In-Context Learning (ICL) has been extensively studied in text-only Language Models, but remains largely unexplored in the speech domain. Here, we inv

researcharxiv-cs-ai
10 Apr 2026
Local Ai

kugel-2 model (VibeVoice finetune) repo is gone. Does anyone know why?

DGX agent

The kugel-2 model is a community fine-tune of Microsoft's VibeVoice, a text-to-speech model — and its disappearance is rooted in Microsoft's own removal of the base VibeVoice repository. Microsoft...

local-air-stablediffusion
10 Apr 2026
Safety

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning

DGX agent

arXiv:2604.07941v1 Announce Type: new Abstract: Post-training has become central to turning pretrained large language models (LLMs) into aligned and deployable systems. Recent progress spans supervise

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

LPM 1.0: Video-based Character Performance Model

DGX agent

arXiv:2604.07823v1 Announce Type: new Abstract: Performance, the externalization of intent, emotion, and personality through visual, vocal, and temporal behavior, is what makes a character alive. Lear

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Non-identifiability of Explanations from Model Behavior in Deep Networks of Image Authenticity Judgments

DGX agent

arXiv:2604.07254v1 Announce Type: cross Abstract: Deep neural networks can predict human judgments, but this does not imply that they rely on human-like information or reveal the cues underlying those

researcharxiv-cs-lg
10 Apr 2026
Research

Operator Learning for Surrogate Modeling of Wave-Induced Forces from Sea Surface Waves

DGX agent

arXiv:2604.06433v1 Announce Type: cross Abstract: Wave setup plays a significant role in transferring wave-induced energy to currents and causing an increase in water elevation. This excess momentum f

researcharxiv-cs-lg
10 Apr 2026
Applications

Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics

DGX agent

arXiv:2604.08503v1 Announce Type: new Abstract: Recent advances in generative video modeling, driven by large-scale datasets and powerful architectures, have yielded remarkable visual realism. However

applicationsarxiv-cs-cv
10 Apr 2026
← Previous
1…120121122123124…1261
Next →