AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,504 results
15 Apr 2026

Conflated Inverse Modeling to Generate Diverse and Temperature-Change Inducing Urban Vegetation Patterns

ResearchDGX agent

arXiv:2604.13028v1 Announce Type: new Abstract: Urban areas are increasingly vulnerable to thermal extremes driven by rapid urbanization and climate change. Traditionally, thermal extremes have been m

Dynamic Modeling and Robust Gait Optimization of a Compliant Worm Robot

ResearchDGX agent

arXiv:2604.12031v1 Announce Type: new Abstract: Worm-inspired robots provide an effective locomotion strategy for constrained environments by combining cyclic body deformation with alternating anchori

GroupKAN: Efficient Kolmogorov-Arnold Networks via Grouped Spline Modeling

Model ReleasesDGX agent

arXiv:2511.05477v2 Announce Type: replace Abstract: Medical image segmentation demands models that achieve high accuracy while maintaining computational efficiency and clinical interpretability. While

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HazardArena: Evaluating Semantic Safety in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.12447v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models inherit rich world knowledge from vision-language backbones and acquire executable skills via action demonstrations.

LLM-HYPER: Generative CTR Modeling for Cold-Start Ad Personalization via LLM-Based Hypernetworks

ApplicationsDGX agent

arXiv:2604.12096v1 Announce Type: new Abstract: On online advertising platforms, newly introduced promotional ads face the cold-start problem, as they lack sufficient user feedback for model training.

MODIX: A Training-Free Multimodal Information-Driven Positional Index Scaling for Vision-Language Models

SafetyDGX agent

arXiv:2604.12537v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in multimodal understanding, yet their positional encoding mechanisms remain suboptima

Narrative over Numbers: The Identifiable Victim Effect and its Amplification Under Alignment and Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.12076v1 Announce Type: cross Abstract: The Identifiable Victim Effect (IVE) - the tendency to allocate greater resources to a specific, narratively described victim than to a statistically

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Professional Image Quality Assessment (Track 1)

Model ReleasesDGX agent

arXiv:2604.12512v1 Announce Type: cross Abstract: In this paper, we present an overview of the NTIRE 2026 challenge on the 3rd Restore Any Image Model in the Wild, specifically focusing on Track 1: Pr

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models

SafetyDGX agent

arXiv:2604.12582v1 Announce Type: new Abstract: Recent Video Large Language Models (Video-LLMs) have demonstrated strong capability in video understanding, yet they still suffer from hallucinations. E

Running a 31B model locally made me realize how insane LLM infra actually is

Local AiDGX agent

A Reddit post from r/ollama in which a user shares their experience running a 31B parameter model locally using Ollama, reflecting on the surprisingly demanding hardware and infrastructure requirement

The example prompt for Google's new Gemini Flash TTS text-to-speed model is a lot https://simonwillison.net/2026/Apr/15/gemini-31-flash-tts/

Model ReleasesDGX agent

Google's Gemini 3.1 Flash TTS (text-to-speech) model includes a notably elaborate or extensive example prompt, which Simon Willison highlighted as noteworthy. The post likely comments on the complexit

Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet. This launch [excitement] includes aud…

Model ReleasesDGX agent

Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet. This launch [excitement] includes audio tags! 🗣🏷 Audio tags [explanatory] are a seamless way to g

Understanding or Memorizing? A Case Study of German Definite Articles in Language Models

Model ReleasesDGX agent

arXiv:2601.09313v2 Announce Type: replace-cross Abstract: Language models perform well on grammatical agreement, but it is unclear whether this reflects rule-based generalization or memorization. We s

Variation in Verification: Understanding Verification Dynamics in Large Language Models

Model ReleasesDGX agent

arXiv:2509.17995v2 Announce Type: replace-cross Abstract: Recent advances have shown that scaling test-time computation enables large language models (LLMs) to solve increasingly complex problems acro

Visual Diffusion Models are Geometric Solvers

ResearchDGX agent

arXiv:2510.21697v2 Announce Type: replace Abstract: In this paper we show that visual diffusion models can serve as effective geometric solvers: they can directly reason about geometric problems by wo

What if you could get 1.3B Transformer quality from a 770M model? That's not a compression result. It's a different architecture. Parcae, fr…

ToolsDGX agent

What if you could get 1.3B Transformer quality from a 770M model? That's not a compression result. It's a different architecture. Parcae, from @realDanFu (Together AI's VP of Kernels) and his lab at U

Which cloud model do you use for coding? Which one got better reasoning?

Local AiDGX agent

This r/ollama thread is a community discussion where users share their preferred cloud-based AI models for coding tasks and compare their reasoning capabilities. The conversation likely highlights pop

Why Did Apple Fall: Evaluating Curiosity in Large Language Models

TutorialsDGX agent

arXiv:2510.20635v2 Announce Type: replace-cross Abstract: Curiosity serves as a pivotal conduit for human beings to discover and learn new knowledge. Recent advancements of large language models (LLMs

14 Apr 2026

BadGraph: A Backdoor Attack Against Latent Diffusion Model for Text-Guided Graph Generation

Model ReleasesDGX agent

arXiv:2510.20792v4 Announce Type: replace-cross Abstract: The rapid progress of graph generation has raised new security concerns, particularly regarding backdoor vulnerabilities. While prior work has

Belief-Aware VLM Model for Human-like Reasoning

SafetyDGX agent

arXiv:2604.09686v1 Announce Type: new Abstract: Traditional neural network models for intent inference rely heavily on observable states and struggle to generalize across diverse tasks and dynamic env

Can Large Language Models Infer Causal Relationships from Real-World Text?

Model ReleasesDGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

Computational Lesions in Multilingual Language Models Separate Shared and Language-specific Brain Alignment

Model ReleasesDGX agent

arXiv:2604.10627v1 Announce Type: cross Abstract: How the brain supports language across different languages is a basic question in neuroscience and a useful test for multilingual artificial intellige

Conflicts Make Large Reasoning Models Vulnerable to Attacks

Model ReleasesDGX agent

arXiv:2604.09750v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have achieved remarkable performance across diverse domains, yet their decision-making under conflicting objectives rema

Cross-Cultural Value Awareness in Large Vision-Language Models

SafetyDGX agent

arXiv:2604.09945v1 Announce Type: cross Abstract: The rapid adoption of large vision-language models (LVLMs) in recent years has been accompanied by growing fairness concerns due to their propensity t

Decoupled Similarity for Task-Aware Token Pruning in Large Vision-Language Models

ResearchDGX agent

arXiv:2604.11240v1 Announce Type: new Abstract: Token pruning has emerged as an effective approach to reduce the substantial computational overhead of Large Vision-Language Models (LVLMs) by discardin

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.03926v2 Announce Type: replace Abstract: The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined poli

EgoFun3D: Modeling Interactive Objects from Egocentric Videos using Function Templates

Model ReleasesDGX agent

arXiv:2604.11038v1 Announce Type: new Abstract: We present EgoFun3D, a coordinated task formulation, dataset, and benchmark for modeling interactive 3D objects from egocentric videos. Interactive obje

Evolutionary Token-Level Prompt Optimization for Diffusion Models

SafetyDGX agent

arXiv:2604.09861v1 Announce Type: new Abstract: Text-to-image diffusion models exhibit strong generative performance but remain highly sensitive to prompt formulation, often requiring extensive manual

General365: Benchmarking General Reasoning in Large Language Models Across Diverse and Challenging Tasks

Model ReleasesDGX agent

arXiv:2604.11778v1 Announce Type: cross Abstract: Contemporary large language models (LLMs) have demonstrated remarkable reasoning capabilities, particularly in specialized domains like mathematics an

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling

Model ReleasesDGX agent

arXiv:2604.07209v2 Announce Type: replace Abstract: Building world models with spatial consistency and real-time interactivity remains a fundamental challenge in computer vision. Current video generat

MorphoFlow: Sparse-Supervised Generative Shape Modeling with Adaptive Latent Relevance

TutorialsDGX agent

arXiv:2604.11636v1 Announce Type: new Abstract: Statistical shape modeling (SSM) is central to population level analysis of anatomical variability, yet most existing approaches rely on densely annotat

New post: We show that small, cheap models can detect the flagship Mythos FreeBSD zero-day (CVE-2026-4747) using a simple harness we call na…

IndustryDGX agent

New post: We show that small, cheap models can detect the flagship Mythos FreeBSD zero-day (CVE-2026-4747) using a simple harness we call nano-analyzer Models down to 3.6B active params (including ope

NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results

Model ReleasesDGX agent

arXiv:2604.10551v1 Announce Type: new Abstract: This paper presents an overview of the NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models. This challenge utili

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: AI Flash Portrait (Track 3)

Model ReleasesDGX agent

arXiv:2604.11230v1 Announce Type: new Abstract: In this paper, we present a comprehensive overview of the NTIRE 2026 3rd Restore Any Image Model (RAIM) challenge, with a specific focus on Track 3: AI

Pay Less Attention to Function Words for Free Robustness of Vision-Language Models

ResearchDGX agent

arXiv:2512.07222v3 Announce Type: replace-cross Abstract: To address the trade-off between robustness and performance for robust VLM, we observe that function words could incur vulnerability of VLMs a

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

Model ReleasesDGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models

Model ReleasesDGX agent

arXiv:2505.15489v4 Announce Type: replace-cross Abstract: The impact of multimodal misinformation arises not only from factual inaccuracies but also from the misleading narratives that creators delibe

Shared Emotion Geometry Across Small Language Models: A Cross-Architecture Study of Representation, Behavior, and Methodological Confounds

Model ReleasesDGX agent

arXiv:2604.11050v1 Announce Type: cross Abstract: We extract 21-emotion vector sets from twelve small language models (six architectures x base/instruct, 1B-8B parameters) under a unified comprehensio

SmileyLlama: Modifying Large Language Models for Directed Chemical Space Exploration

Model ReleasesDGX agent

arXiv:2409.02231v5 Announce Type: replace-cross Abstract: We show that large language model (LLMs) can be transformed via supervised fine-tuning (SFT) of engineered prompts into SmileyLlama for explor

The Weight of a Bit: EMFI Sensitivity Analysis of Embedded Deep Learning Models

ResearchDGX agent

arXiv:2602.16309v2 Announce Type: replace-cross Abstract: Fault injection attacks on embedded neural network models have been shown as a potent threat. Numerous works studied resilience of models from

Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities

Model ReleasesDGX agent

arXiv:2604.10135v1 Announce Type: cross Abstract: Researchers have explored different ways to improve large language models (LLMs)' capabilities via dummy token insertion in contexts. However, existin

VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation

Model ReleasesDGX agent

arXiv:2604.10127v1 Announce Type: cross Abstract: The rapid advancement of AIGC-based video generation has underscored the critical need for comprehensive evaluation frameworks that go beyond traditio

Why Do Multilingual Reasoning Gaps Emerge in Reasoning Language Models?

ResearchDGX agent

arXiv:2510.27269v3 Announce Type: replace-cross Abstract: Reasoning language models (RLMs) achieve strong performance on complex reasoning tasks, yet they still exhibit a multilingual reasoning gap, p

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

Model ReleasesDGX agent

arXiv:2602.02343v3 Announce Type: replace-cross Abstract: Methods for controlling large language models (LLMs), including local weight fine-tuning, LoRA-based adaptation, and activation-based interven

13 Apr 2026

Adaptive Action Chunking at Inference-time for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2604.04161v2 Announce Type: replace Abstract: In Vision-Language-Action (VLA) models, action chunking (i.e., executing a sequence of actions without intermediate replanning) is a key technique t

AssemLM: Spatial Reasoning Multimodal Large Language Models for Robotic Assembly

Model ReleasesDGX agent

arXiv:2604.08983v1 Announce Type: new Abstract: Spatial reasoning is a fundamental capability for embodied intelligence, especially for fine-grained manipulation tasks such as robotic assembly. While

AudioGuard: Toward Comprehensive Audio Safety Protection Across Diverse Threat Models

Model ReleasesDGX agent

arXiv:2604.08867v1 Announce Type: cross Abstract: Audio has rapidly become a primary interface for foundation models, powering real-time voice assistants. Ensuring safety in audio systems is inherentl

BlendFusion -- Scalable Synthetic Data Generation for Diffusion Model Training

ResearchDGX agent

arXiv:2604.09022v1 Announce Type: new Abstract: With the rapid adoption of diffusion models, synthetic data generation has emerged as a promising approach for addressing the growing demand for large-s

Cards Against LLMs: Benchmarking Humor Alignment in Large Language Models

SafetyDGX agent

arXiv:2604.08757v1 Announce Type: cross Abstract: Humor is one of the most culturally embedded and socially significant dimensions of human communication, yet it remains largely unexplored as a dimens

Decomposing the Delta: What Do Models Actually Learn from Preference Pairs?

TutorialsDGX agent

arXiv:2604.08723v1 Announce Type: cross Abstract: Preference optimization methods such as DPO and KTO are widely used for aligning language models, yet little is understood about what properties of pr

EvoLen: Evolution-Guided Tokenization for DNA Language Model

SafetyDGX agent

arXiv:2604.08698v1 Announce Type: new Abstract: Tokens serve as the basic units of representation in DNA language models (DNALMs), yet their design remains underexplored. Unlike natural language, DNA

From Navigation to Refinement: Revealing the Two-Stage Nature of Flow-based Diffusion Models through Oracle Velocity

ResearchDGX agent

arXiv:2512.02826v3 Announce Type: replace-cross Abstract: Flow-based diffusion models have emerged as a leading paradigm for training generative models across images and videos. However, their memoriz

HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models

ResearchDGX agent

arXiv:2604.06165v2 Announce Type: replace Abstract: Large vision-language models can produce object hallucinations in image descriptions, highlighting the need for effective detection and mitigation s

Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling

Model ReleasesDGX agent

arXiv:2604.05072v2 Announce Type: replace Abstract: Recent large language models have shifted SVG generation from differentiable rendering optimization to autoregressive program synthesis. However, ex

Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models

Model ReleasesDGX agent

arXiv:2604.08970v1 Announce Type: cross Abstract: We study predictive multilingual evaluation: estimating how well a model will perform on a task in a target language when direct benchmark results are

Multi-task Just Recognizable Difference for Video Coding for Machines: Database, Model, and Coding Application

ResearchDGX agent

arXiv:2604.09421v1 Announce Type: cross Abstract: Just Recognizable Difference (JRD) boosts coding efficiency for machine vision through visibility threshold modeling, but is currently limited to a si

Post-Hoc Guidance for Consistency Models by Joint Flow Distribution Learning

SafetyDGX agent

arXiv:2604.08828v1 Announce Type: cross Abstract: Classifier-free Guidance (CFG) lets practitioners trade-off fidelity against diversity in Diffusion Models (DMs). The practicality of CFG is however h

Predictive Entropy Links Calibration and Paraphrase Sensitivity in Medical Vision-Language Models

ResearchDGX agent

arXiv:2604.08941v1 Announce Type: new Abstract: Medical Vision Language Models VLMs suffer from two failure modes that threaten safe deployment mis calibrated confidence and sensitivity to question re

Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models

Local AiDGX agent

arXiv:2505.12509v3 Announce Type: replace-cross Abstract: Post-hoc explanations provide transparency and are essential for guiding model optimization, such as prompt engineering and data sanitation. H

← Previous
1…9596979899…1009
Next →