AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,537 results
Model Releases

Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities

DGX agent

arXiv:2604.10135v1 Announce Type: cross Abstract: Researchers have explored different ways to improve large language models (LLMs)' capabilities via dummy token insertion in contexts. However, existin

model-releasesarxiv-cs-ai
14 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation

DGX agent

arXiv:2604.10127v1 Announce Type: cross Abstract: The rapid advancement of AIGC-based video generation has underscored the critical need for comprehensive evaluation frameworks that go beyond traditio

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Why Do Multilingual Reasoning Gaps Emerge in Reasoning Language Models?

DGX agent

arXiv:2510.27269v3 Announce Type: replace-cross Abstract: Reasoning language models (RLMs) achieve strong performance on complex reasoning tasks, yet they still exhibit a multilingual reasoning gap, p

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

DGX agent

arXiv:2602.02343v3 Announce Type: replace-cross Abstract: Methods for controlling large language models (LLMs), including local weight fine-tuning, LoRA-based adaptation, and activation-based interven

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

Adaptive Action Chunking at Inference-time for Vision-Language-Action Models

DGX agent

arXiv:2604.04161v2 Announce Type: replace Abstract: In Vision-Language-Action (VLA) models, action chunking (i.e., executing a sequence of actions without intermediate replanning) is a key technique t

applicationsarxiv-cs-ro
13 Apr 2026
Model Releases

AssemLM: Spatial Reasoning Multimodal Large Language Models for Robotic Assembly

DGX agent

arXiv:2604.08983v1 Announce Type: new Abstract: Spatial reasoning is a fundamental capability for embodied intelligence, especially for fine-grained manipulation tasks such as robotic assembly. While

model-releasesarxiv-cs-ro
13 Apr 2026
Model Releases

AudioGuard: Toward Comprehensive Audio Safety Protection Across Diverse Threat Models

DGX agent

arXiv:2604.08867v1 Announce Type: cross Abstract: Audio has rapidly become a primary interface for foundation models, powering real-time voice assistants. Ensuring safety in audio systems is inherentl

model-releasesarxiv-cs-ai
13 Apr 2026
Research

BlendFusion -- Scalable Synthetic Data Generation for Diffusion Model Training

DGX agent

arXiv:2604.09022v1 Announce Type: new Abstract: With the rapid adoption of diffusion models, synthetic data generation has emerged as a promising approach for addressing the growing demand for large-s

researcharxiv-cs-cv
13 Apr 2026
Safety

Cards Against LLMs: Benchmarking Humor Alignment in Large Language Models

DGX agent

arXiv:2604.08757v1 Announce Type: cross Abstract: Humor is one of the most culturally embedded and socially significant dimensions of human communication, yet it remains largely unexplored as a dimens

safetyarxiv-cs-ai
13 Apr 2026
Tutorials

Decomposing the Delta: What Do Models Actually Learn from Preference Pairs?

DGX agent

arXiv:2604.08723v1 Announce Type: cross Abstract: Preference optimization methods such as DPO and KTO are widely used for aligning language models, yet little is understood about what properties of pr

tutorialsarxiv-cs-ai
13 Apr 2026
Safety

EvoLen: Evolution-Guided Tokenization for DNA Language Model

DGX agent

arXiv:2604.08698v1 Announce Type: new Abstract: Tokens serve as the basic units of representation in DNA language models (DNALMs), yet their design remains underexplored. Unlike natural language, DNA

safetyarxiv-cs-lg
13 Apr 2026
Research

From Navigation to Refinement: Revealing the Two-Stage Nature of Flow-based Diffusion Models through Oracle Velocity

DGX agent

arXiv:2512.02826v3 Announce Type: replace-cross Abstract: Flow-based diffusion models have emerged as a leading paradigm for training generative models across images and videos. However, their memoriz

researcharxiv-cs-ai
13 Apr 2026
Research

HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models

DGX agent

arXiv:2604.06165v2 Announce Type: replace Abstract: Large vision-language models can produce object hallucinations in image descriptions, highlighting the need for effective detection and mitigation s

researcharxiv-cs-cv
13 Apr 2026
Model Releases

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling

DGX agent

arXiv:2604.05072v2 Announce Type: replace Abstract: Recent large language models have shifted SVG generation from differentiable rendering optimization to autoregressive program synthesis. However, ex

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models

DGX agent

arXiv:2604.08970v1 Announce Type: cross Abstract: We study predictive multilingual evaluation: estimating how well a model will perform on a task in a target language when direct benchmark results are

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Multi-task Just Recognizable Difference for Video Coding for Machines: Database, Model, and Coding Application

DGX agent

arXiv:2604.09421v1 Announce Type: cross Abstract: Just Recognizable Difference (JRD) boosts coding efficiency for machine vision through visibility threshold modeling, but is currently limited to a si

researcharxiv-cs-cv
13 Apr 2026
Safety

Post-Hoc Guidance for Consistency Models by Joint Flow Distribution Learning

DGX agent

arXiv:2604.08828v1 Announce Type: cross Abstract: Classifier-free Guidance (CFG) lets practitioners trade-off fidelity against diversity in Diffusion Models (DMs). The practicality of CFG is however h

safetyarxiv-cs-cv
13 Apr 2026
Research

Predictive Entropy Links Calibration and Paraphrase Sensitivity in Medical Vision-Language Models

DGX agent

arXiv:2604.08941v1 Announce Type: new Abstract: Medical Vision Language Models VLMs suffer from two failure modes that threaten safe deployment mis calibrated confidence and sensitivity to question re

researcharxiv-cs-lg
13 Apr 2026
Local Ai

Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models

DGX agent

arXiv:2505.12509v3 Announce Type: replace-cross Abstract: Post-hoc explanations provide transparency and are essential for guiding model optimization, such as prompt engineering and data sanitation. H

local-aiarxiv-cs-ai
13 Apr 2026
Research

State Space Models are Effective Sign Language Learners: Exploiting Phonological Compositionality for Vocabulary-Scale Recognition

DGX agent

arXiv:2604.08761v1 Announce Type: new Abstract: Sign language recognition suffers from catastrophic scaling failure: models achieving high accuracy on small vocabularies collapse at realistic sizes. E

researcharxiv-cs-cv
13 Apr 2026
Model Releases

The AI Codebase Maturity Model: From Assisted Coding to Self-Sustaining Systems

DGX agent

arXiv:2604.09388v1 Announce Type: cross Abstract: AI coding tools are widely adopted, but most teams plateau at prompt-and-review without a framework for systematic progression. This paper presents th

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Toward World Models for Epidemiology

DGX agent

arXiv:2604.09519v1 Announce Type: new Abstract: World models have emerged as a unifying paradigm for learning latent dynamics, simulating counterfactual futures, and supporting planning under uncertai

safetyarxiv-cs-lg
13 Apr 2026
Tutorials

Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models

DGX agent

arXiv:2604.09227v1 Announce Type: cross Abstract: Image generative models have become indispensable tools to yield exquisite high-resolution (HR) images for everyone, ranging from general users to pro

tutorialsarxiv-cs-cv
13 Apr 2026
Local Ai

Tried Ollama Cloud, just realize only Kimi model accept images

DGX agent

A Reddit user exploring Ollama Cloud noted that, at the time of their post, only the Kimi model supported image (vision/multimodal) inputs among the available cloud models. Kimi K2.5 is a native multi

local-air-ollama
13 Apr 2026
Research

Uncertainty-Aware Transformers: Conformal Prediction for Language Models

DGX agent

arXiv:2604.08885v1 Announce Type: new Abstract: Transformers have had a profound impact on the field of artificial intelligence, especially on large language models and their variants. However, as was

researcharxiv-cs-lg
13 Apr 2026
Model Releases

We conducted cyber evaluations of Claude Mythos Preview and found that it is the first model to complete an AISI cyber range end-to-end. 🧵

DGX agent

Boris Cherny and colleagues conducted cybersecurity evaluations of Claude Mythos Preview, finding it to be the first AI model to successfully complete an AISI (AI Safety Institute) cyber range end-to-

model-releasesboris-cherny--x
13 Apr 2026
Tutorials

How to train loras in One Trainer for Z Image using Civitai models?

DGX agent

This r/StableDiffusion post likely discusses the community challenge of using OneTrainer — a locally-installed, GUI-based LoRA training tool — to train LoRAs specifically compatible with Z Image (Tong

tutorialsr-stablediffusion
12 Apr 2026
Local Ai

@hwchase17 I think harness/managed agents is a way for Anthropic to keep its Moat. As models get mature, the need for cloud LLMs might reduc…

DGX agent

@hwchase17 I think harness/managed agents is a way for Anthropic to keep its Moat. As models get mature, the need for cloud LLMs might reduce and local LLM models might increase (save cost) and thus t

local-aiharrison-chase--x
12 Apr 2026
Model Releases

KIV: 1M token context window on a RTX 4070 (12GB VRAM), no retraining, drop-in HuggingFace cache replacement - Works with any model that uses DynamicCache [P]

DGX agent

KIV is a project shared on r/MachineLearning presenting a drop-in replacement for HuggingFace's `DynamicCache` that enables up to 1 million token context windows on consumer hardware with only 12GB of

model-releasesr-machinelearning
12 Apr 2026
Local Ai

What's model should I run?

DGX agent

A Reddit discussion from the r/ollama community where a user seeks advice on which AI language model to run locally using Ollama. Responses likely include hardware-based recommendations (such as RAM a

local-air-ollama
11 Apr 2026
Model Releases

A Benchmark of Classical and Deep Learning Models for Agricultural Commodity Price Forecasting on A Novel Bangladeshi Market Price Dataset

DGX agent

arXiv:2604.06227v1 Announce Type: new Abstract: Accurate short-term forecasting of agricultural commodity prices is critical for food security planning and smallholder income stabilisation in developi

model-releasesarxiv-cs-lg
10 Apr 2026
Applications

A Comparative Study of Demonstration Selection for Practical Large Language Models-based Next POI Prediction

DGX agent

arXiv:2604.06207v1 Announce Type: cross Abstract: This paper investigates demonstration selection strategies for predicting a user's next point-of-interest (POI) using large language models (LLMs), ai

applicationsarxiv-cs-ai
10 Apr 2026
Model Releases

AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors

DGX agent

arXiv:2601.20524v2 Announce Type: replace Abstract: Zero-shot anomaly detection aims to detect and localise abnormal regions in the image without access to any in-domain training images. While recent

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models

DGX agent

arXiv:2509.23435v2 Announce Type: replace-cross Abstract: The creation of high-quality multimodal datasets remains fundamental for advancing role-playing capabilities in large language models (LLMs).

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Bootstrapping Sign Language Annotations with Sign Language Models

DGX agent

arXiv:2604.07606v1 Announce Type: new Abstract: AI-driven sign language interpretation is limited by a lack of high-quality annotated data. New datasets including ASL STEM Wiki and FLEURS-ASL contain

model-releasesarxiv-cs-cv
10 Apr 2026
Tools

ChatGPT voice mode is a weaker model

DGX agent

I think it's non-obvious to many people that the OpenAI voice mode runs on a much older, much weaker model - it feels like the AI that you can talk to should be the smartest AI but it really isn't. If

toolssimon-willison
10 Apr 2026
Research

Compact Example-Based Explanations for Language Models

DGX agent

arXiv:2601.03786v2 Announce Type: replace Abstract: Training data influence estimation methods quantify the contribution of training documents to a model's output, making them a promising source of in

researcharxiv-cs-cl
10 Apr 2026
Local Ai

ConfusionPrompt: Practical Private Inference for Online Large Language Models

DGX agent

arXiv:2401.00870v5 Announce Type: replace-cross Abstract: State-of-the-art large language models (LLMs) are typically deployed as online services, requiring users to transmit detailed prompts to cloud

local-aiarxiv-cs-ai
10 Apr 2026
Model Releases

Detecting HIV-Related Stigma in Clinical Narratives Using Large Language Models

DGX agent

arXiv:2604.07717v1 Announce Type: new Abstract: Human immunodeficiency virus (HIV)-related stigma is a critical psychosocial determinant of health for people living with HIV (PLWH), influencing mental

model-releasesarxiv-cs-cl
10 Apr 2026
Research

DiffVC: A Non-autoregressive Framework Based on Diffusion Model for Video Captioning

DGX agent

arXiv:2604.08084v1 Announce Type: new Abstract: Current video captioning methods usually use an encoder-decoder structure to generate text autoregressively. However, autoregressive methods have inhere

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models

DGX agent

arXiv:2604.08284v1 Announce Type: new Abstract: Large language models store not only isolated facts but also rules that support reasoning across symbolic expressions, natural language explanations, an

model-releasesarxiv-cs-cl
10 Apr 2026
Applications

Environmental, Social and Governance Sentiment Analysis on Slovene News: A Novel Dataset and Models

DGX agent

arXiv:2604.06826v1 Announce Type: cross Abstract: Environmental, Social, and Governance (ESG) considerations are increasingly integral to assessing corporate performance, reputation, and long-term sus

applicationsarxiv-cs-ai
10 Apr 2026
Agents

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexib…

DGX agent

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexibility, but to ensure we actually own the long-term memory th

agentsharrison-chase--x
10 Apr 2026
Applications

HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models

DGX agent

arXiv:2512.09928v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently enabled robotic manipulation by grounding visual and linguistic cues into actions. However, most V

applicationsarxiv-cs-ro
10 Apr 2026
Tools

If you ask ChatGPT voice mode for its knowledge cutoff date it tells you April 2024 - it's a GPT-4o era model

DGX agent

ChatGPT's voice mode, when directly queried about its knowledge cutoff date, reports April 2024, indicating it is powered by a GPT-4o era model rather than a more recent one. GPT-4o initially had ...

toolssimon-willison--x
10 Apr 2026
Research

In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads

DGX agent

arXiv:2604.06356v1 Announce Type: cross Abstract: In-Context Learning (ICL) has been extensively studied in text-only Language Models, but remains largely unexplored in the speech domain. Here, we inv

researcharxiv-cs-ai
10 Apr 2026
Local Ai

kugel-2 model (VibeVoice finetune) repo is gone. Does anyone know why?

DGX agent

The kugel-2 model is a community fine-tune of Microsoft's VibeVoice, a text-to-speech model — and its disappearance is rooted in Microsoft's own removal of the base VibeVoice repository. Microsoft...

local-air-stablediffusion
10 Apr 2026
Safety

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning

DGX agent

arXiv:2604.07941v1 Announce Type: new Abstract: Post-training has become central to turning pretrained large language models (LLMs) into aligned and deployable systems. Recent progress spans supervise

safetyarxiv-cs-cl
10 Apr 2026
← Previous
1…120121122123124…1262
Next →