AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,036 results
Local Ai

EasyLens: A Training-Free Plug-and-Play Subtle-Lesion Representation Amplifier for Medical Vision-Language Models

DGX agent

arXiv:2606.06379v1 Announce Type: new Abstract: Medical vision-language models (VLMs) have shown increasing potential for clinical image interpretation, including lesion detection and report generatio

local-aiarxiv-cs-cv
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

FontFusion: Enhancing Generative Text in Diffusion Models with Typographic Conditioning

DGX agent

arXiv:2606.06066v1 Announce Type: new Abstract: Typography generation in diffusion models faces a persistent trade-off: enabling precise font control typically degrades text legibility, while maintain

researcharxiv-cs-cv
5 Jun 2026
Research

Framing, Judging, Steering: An Assessable Competency Model for Teach-ing Students to Reason With Generative AI

DGX agent

arXiv:2606.05983v1 Announce Type: cross Abstract: Generative AI makes answers easy and understanding hard, and uncritical use invites cognitive offloading. Schools still measure unaided performance, y

researcharxiv-cs-cl
5 Jun 2026
Safety

Geometry-Aware Dataset Condensation for Diffusion Model Training

DGX agent

arXiv:2606.05883v1 Announce Type: new Abstract: Dataset condensation aims to construct compact datasets from real data via synthesis or selection. However, existing approaches are ill-suited for diffu

safetyarxiv-cs-cv
5 Jun 2026
Research

IR3DE: A Linear Router for Large Language Models

DGX agent

arXiv:2606.06098v1 Announce Type: new Abstract: Foundational Large Language Models (LLMs) demonstrate proficiency on a wide range of general tasks, and achieve remarkable results on various specialize

researcharxiv-cs-cl
5 Jun 2026
Safety

Large Language Models are Perplexed by some Political Parties

DGX agent

arXiv:2606.05937v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used, including in political applications, but their political fairness has been little studied. We assess

safetyarxiv-cs-cl
5 Jun 2026
Local Ai

NAVIRA: Decoupled Stochastic Remasking for Masked Diffusion Language Models

DGX agent

arXiv:2606.06031v1 Announce Type: new Abstract: Masked diffusion language models generate text by iteratively unmasking many tokens in parallel, but this speed comes with a correction problem: tokens

local-aiarxiv-cs-cl
5 Jun 2026
Agents

OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions

DGX agent

arXiv:2602.05843v2 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) has catalyzed the development of autonomous agents capable of navigating complex environments.

agentsarxiv-cs-cl
5 Jun 2026
Safety

PiL-World: A Chunk-Wise World Model for VLA Policy-in-the-Loop Evaluation

DGX agent

arXiv:2606.05773v1 Announce Type: new Abstract: Vision-language-action (VLA) policies operate in a closed loop in real-world robot tasks: a robot observes the scene, executes an action chunk, and cond

safetyarxiv-cs-ro
5 Jun 2026
Research

ReTreVal: Reasoning Tree with Validation and Cross-Problem Memory for Large Language Models

DGX agent

arXiv:2601.02880v2 Announce Type: replace-cross Abstract: Every existing inference-time reasoning framework discards all failure context at problem boundaries, leaving a model solving problem 500 no w

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Scaffold, Not Vocabulary? A Controlled, Two-Tier, Pre-Registered Study of a Popperian Code-Generation Skill

DGX agent

arXiv:2606.06454v1 Announce Type: cross Abstract: Large language models increasingly write, review, and judge code, and a fast-growing practice equips them with prompt 'skills' that ask the model to r

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

UNIVID: Unified Vision-Language Model for Video Moderation

DGX agent

arXiv:2606.05748v1 Announce Type: cross Abstract: Global-scale video moderation faces a dual challenge: the need for fine-grained multi-modal reasoning and the demand for interpretable outputs to supp

safetyarxiv-cs-cl
5 Jun 2026
Research

Best Visual Reasoning Model in 2026 (Including APIs) [D]

DGX agent

Gemini 3.1 Pro and Gemini 3-Pro lead visual reasoning benchmarks , with GPT-5.2, Kimi-K2.5, and GPT-5.2-Pro following . A 2026 evaluation benchmarked 15 leading multimodal models on visual reasoning a

researchr-machinelearning
4 Jun 2026
Research

Beyond Text Following: Repairable Arbitration Reversals in Audio-Language Models

DGX agent

arXiv:2606.05161v1 Announce Type: cross Abstract: Audio-language models (ALMs) often follow text that conflicts with audio, even when the audio evidence is clear. This raises a basic question: is the

researcharxiv-cs-cl
4 Jun 2026
Safety

Causal Multi-fidelity Surrogate Forward and Inverse Models for ICF Implosions

DGX agent

arXiv:2509.05510v3 Announce Type: replace-cross Abstract: Continued progress in inertial confinement fusion (ICF) requires solving inverse problems relating experimental observations to simulation inp

safetyarxiv-cs-lg
4 Jun 2026
Safety

Covert Influence Between Language Models

DGX agent

arXiv:2606.04071v1 Announce Type: cross Abstract: As language models increasingly consume one another's outputs, covert influence -- a phenomenon where a sender's payload (the behavioral disposition i

safetyarxiv-cs-cl
4 Jun 2026
Safety

Efficient Adversarial Attacks on High-dimensional Offline Bandits

DGX agent

arXiv:2602.01658v2 Announce Type: replace-cross Abstract: Bandit algorithms have recently emerged as a powerful tool for evaluating machine learning models, including generative image models and large

safetyarxiv-cs-ai
4 Jun 2026
Research

Efficient and Training-Free Single-Image Diffusion Models

DGX agent

arXiv:2606.04299v1 Announce Type: new Abstract: We consider the problem of generating images whose internal structure -- defined by the distribution of patches across multiple scales -- matches that o

researcharxiv-cs-cv
4 Jun 2026
Safety

Geospatial Foundation Models to Enable Progress on Sustainable Development Goals

DGX agent

arXiv:2505.24528v3 Announce Type: replace Abstract: Foundation Models (FMs) are large-scale, pre-trained artificial intelligence (AI) systems that have revolutionized natural language processing and c

safetyarxiv-cs-cv
4 Jun 2026
Tutorials

Learning What to Learn: Stage-Specific Data Sets for SFT-then-RL in Small Language Model Reasoning

DGX agent

arXiv:2606.04466v1 Announce Type: new Abstract: Post-training Small Language Models (SLMs) for reasoning typically follows an SFT-then-RL pipeline, yet existing work rarely considers what data should

tutorialsarxiv-cs-cl
4 Jun 2026
Local Ai

Meet LM Studio's mobile app. Your local models, now in your pocket.

DGX agent

LM Studio has released a mobile app that enables users to run and access local language models on their smartphones. The app extends LM Studio's desktop functionality, allowing users to deploy and int

local-ailm-studio--x
4 Jun 2026
Safety

MorphoQuant: Modality-Aware Quantization for Omni-modal Large Language Models

DGX agent

arXiv:2606.04349v1 Announce Type: cross Abstract: Conventional Post-Training Quantization (PTQ) methods struggle with 4-bit Omni-modal Large Language Models (OLLMs) due to the extreme distribution het

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --mo…

DGX agent

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --model gemma4:12b Hermes Agent ollama launch hermes --model gem

model-releasesollama--x
4 Jun 2026
Research

Physics-Informed Neural Engine Sound Modeling with Differentiable Pulse-Train Synthesis

DGX agent

arXiv:2603.09391v2 Announce Type: replace-cross Abstract: Engine sounds originate from sequential exhaust pressure pulses rather than sustained harmonic oscillations. While neural synthesis methods ty

researcharxiv-cs-ai
4 Jun 2026
Research

SFMP: Fine-Grained, Hardware-Friendly and Search-Free Mixed-Precision Quantization for Large Language Models

DGX agent

arXiv:2602.01027v2 Announce Type: replace Abstract: Mixed-precision quantization is a promising approach for compressing large language models under tight memory budgets. However, existing mixed-preci

researcharxiv-cs-lg
4 Jun 2026
Research

Spatially Grounded Concept Bottleneck Models via Part-Factorized Attention

DGX agent

arXiv:2606.04364v1 Announce Type: new Abstract: Concept bottleneck models (CBMs) predict a layer of human-named attributes before predicting a class, which makes their decisions auditable. On fine-gra

researcharxiv-cs-cv
4 Jun 2026
Research

STaR-Quant: State-Time Consistent Post-Training Quantization for Diffusion Large Language Models

DGX agent

arXiv:2606.04945v1 Announce Type: new Abstract: Diffusion large language models (DLLMs) have recently emerged as a promising alternative to autoregressive LLMs by generating text through iterative mas

researcharxiv-cs-lg
4 Jun 2026
Tutorials

SurvPFN: Towards Foundation Models for Survival Predictions

DGX agent

arXiv:2606.04564v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have made rapid progress in standard classification and regression, but time-to-event survival prediction tasks have re

tutorialsarxiv-cs-lg
4 Jun 2026
Safety

Test-time reward-guided alignment of language models by importance sampling on pre-logit space

DGX agent

arXiv:2510.26219v3 Announce Type: replace-cross Abstract: Test-time alignment of large language models (LLMs) attracts attention because fine-tuning of LLMs requires high computational costs. In this

safetyarxiv-cs-ai
4 Jun 2026
Local Ai

Towards Estimating Normal and Shear Interface Pressures in Prosthetic Sockets via Least Squares and Mechanics Modeling

DGX agent

arXiv:2606.04222v1 Announce Type: new Abstract: Prosthetic socket fitting remains largely manual and iterative, and objective fit metrics are still limited. Part of the challenge is the lack of long-t

local-aiarxiv-cs-ro
4 Jun 2026
Safety

Transferable Multi-Bit Watermarking Across Frozen Diffusion Models via Latent Consistency Bridges

DGX agent

arXiv:2603.20304v2 Announce Type: replace Abstract: As generative AI advances, global governance frameworks increasingly mandate verifiable content provenance. However, existing watermarking technique

safetyarxiv-cs-cv
4 Jun 2026
Research

Who Needs Labels? Adapting Vision Foundation Models With the Metadata You Already Have

DGX agent

arXiv:2606.05107v1 Announce Type: cross Abstract: We propose a label-free approach to adapt powerful but generic vision foundation models to specialized scientific domains. Standard supervised fine-tu

researcharxiv-cs-ai
4 Jun 2026
Local Ai

A Graph Foundation Model with Spectral Parsing and Prototype-Guided Spatial Propagation

DGX agent

arXiv:2606.03315v1 Announce Type: new Abstract: Graph foundation models aim to learn transferable knowledge from diverse graphs for generalization to unseen graphs and tasks. Unlike text and images, g

local-aiarxiv-cs-lg
3 Jun 2026
Safety

A Negative Result on Cross-Model Activation Transfer in a Pythia Multi-Hop Setting

DGX agent

arXiv:2606.03280v1 Announce Type: new Abstract: Recent work shows that language models can transmit behavioural traits through hidden signals in generated data during training. We ask whether a more d

safetyarxiv-cs-ai
3 Jun 2026
Safety

A Pocket Offline Model for Simultaneous Speech Translation as CUNI Submission to IWSLT 2026

DGX agent

arXiv:2606.03948v1 Announce Type: new Abstract: We implement simultaneous translation capability with the offline direct speech-to-text translation model Canary, using the state-of-the-art policy Alig

safetyarxiv-cs-cl
3 Jun 2026
Research

Beyond False Stability: High-Noise Drift Gating for Test-Time Adversarial Defenses in Vision-Language Models

DGX agent

arXiv:2606.03730v1 Announce Type: new Abstract: Vision-language models (VLMs) such as CLIP show strong zero-shot generalization but remain highly vulnerable to adversarial attacks. Adversarial trainin

researcharxiv-cs-cv
3 Jun 2026
Research

Code-on-Graph: Iterative Programmatic Reasoning via Large Language Models on Knowledge Graphs

DGX agent

arXiv:2606.03705v1 Announce Type: new Abstract: Knowledge Graphs (KGs) are widely used to mitigate the limitations of Large Language Models (LLMs), such as outdated knowledge and hallucinations. Exist

researcharxiv-cs-ai
3 Jun 2026
Research

Conformal Language Modeling via Posterior Sampling

DGX agent

arXiv:2606.03731v1 Announce Type: new Abstract: Large Language Models remain plagued by hallucinations. Recent work has sought to tame their prevalence using statistical techniques based on conformal

researcharxiv-cs-lg
3 Jun 2026
Applications

Does Language Shift Break Medical Vision-Language Models? Indonesian Radiology Visual Question Answering Case Study

DGX agent

arXiv:2606.03693v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) are typically evaluated on English radiology visual question answering benchmarks, leaving their robustness under

applicationsarxiv-cs-cl
3 Jun 2026
Safety

Exploring Adversarial Robustness and Safety Alignment in Multilingual Multi-Modal Large Language Models

DGX agent

arXiv:2606.03793v1 Announce Type: new Abstract: Multimodal Large Language Models integrate visual perception into language reasoning, introducing a continuous attack surface susceptible to adversarial

safetyarxiv-cs-cl
3 Jun 2026
Applications

Hybrid Autoregressive-Diffusion Model for Real-Time Sign Language Production

DGX agent

arXiv:2507.09105v4 Announce Type: replace Abstract: Earlier Sign Language Production (SLP) models typically relied on autoregressive decoding, which naturally preserves temporal causality but suffers

applicationsarxiv-cs-cv
3 Jun 2026
Research

Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models

DGX agent

arXiv:2606.03988v1 Announce Type: new Abstract: Vision language models (VLMs) excel at many tasks but still struggle with spatial reasoning when critical information is not directly observable. Many s

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Knowledge-Preserved Model Tuning in Null-Space for Robust Spatio-Temporal Video Grounding

DGX agent

arXiv:2606.03539v1 Announce Type: new Abstract: Spatio-Temporal Video Grounding aims to localize object tubes based on textual queries. While recent methods have achieved remarkable success, they main

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

LAMP: Data-Efficient Linear Affine Weight-Space Models for Parameter-Controlled 3D Shape Generation and Extrapolation

DGX agent

arXiv:2510.22491v3 Announce Type: replace-cross Abstract: Generating high-fidelity 3D geometries under explicit parameter constraints is central to engineering design, yet current methods often requir

model-releasesarxiv-cs-cv
3 Jun 2026
Safety

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models

DGX agent

arXiv:2512.05530v2 Announce Type: replace Abstract: Recently, multimodal large language models (MLLMs) have been widely applied to reasoning tasks. However, they suffer from limited multi-rationale se

safetyarxiv-cs-ai
3 Jun 2026
Safety

Multi-component Causal Tracing in Large Language Models

DGX agent

arXiv:2606.03085v1 Announce Type: cross Abstract: Causal tracing systematically intervenes on a large language model's (LLM's) internal representations to uncover and quantify the causal pathways link

safetyarxiv-cs-cl
3 Jun 2026
Safety

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

DGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

safetyarxiv-cs-cl
3 Jun 2026
Local Ai

ParaBlock: Communication-Computation Parallel Block Coordinate Federated Learning for Large Language Models

DGX agent

arXiv:2511.19959v2 Announce Type: replace Abstract: Federated learning (FL) has been extensively studied as a privacy-preserving training paradigm. Recently, federated block coordinate descent scheme

local-aiarxiv-cs-lg
3 Jun 2026
← Previous
1…226227228229230…1272
Next →