AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

Reinforcement learning to improve large language model-based automated code compliance systems

DGX agent

arXiv:2606.22402v1 Announce Type: cross Abstract: Large language model (LLM)-based approaches for automated code compliance (ACC) of building regulations are prone to generating incorrect and hallucin

model-releasesarxiv-cs-lg
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation

DGX agent

arXiv:2511.20889v2 Announce Type: replace Abstract: Test-time alignment (TTA) aims to adapt models to specific rewards during inference. However, existing methods tend to either under-optimise or over

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

Zero-Shot Vision-Language Models for Classroom Engagement Recognition: A Benchmark Study of Prompt Sensitivity and Cross-Dataset Generalization

DGX agent

arXiv:2606.21861v1 Announce Type: new Abstract: Automated classroom engagement recognition holds substantial promise for scalable learning analytics, yet the suitability of modern Vision-Language Mode

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Models

DGX agent

arXiv:2606.12203v1 Announce Type: new Abstract: Large language models (LLMs) are widely used to tackle complex tasks with autonomous workflows. Recently, reusable natural language skills have emerged

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Bridging the Morphology Gap: Adapting VLA Models to Dexterous Manipulation via Intent-Conditioned Fine-Tuning

DGX agent

arXiv:2606.12109v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable zero-shot generalization in robotic manipulation, yet the vast majority of pre-traine

model-releasesarxiv-cs-ai
11 Jun 2026
Research

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards

DGX agent

arXiv:2511.16672v4 Announce Type: replace Abstract: Recent advances in large multimodal models (LMMs) have enabled impressive reasoning and perception abilities, yet most existing training pipelines s

researcharxiv-cs-cv
11 Jun 2026
Model Releases

PermDoRA -- Understanding Adapter Interference in Language Models: Limits of Parameter-Space Geometry

DGX agent

arXiv:2606.11262v1 Announce Type: cross Abstract: Access control in large language models (LLMs) requires modular mechanisms to enable domain-specific behavior without retraining or cross-domain inter

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models

DGX agent

arXiv:2606.12412v1 Announce Type: cross Abstract: Vision-language models (VLMs) project images into hundreds to thousands of visual tokens, making decoder inference expensive in both attention computa

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Sparse probes and murky physics: a case study of interpretability challenges in a foundation model for continuum dynamics

DGX agent

arXiv:2606.11657v1 Announce Type: cross Abstract: Generative AI emulators are increasingly used in scientific domains where we already have strong theory, benchmarks, and physical intuition. This rais

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Task-Aligned Stability Analysis of Vision-Language Models for Autonomous Driving Hazard Detection

DGX agent

arXiv:2606.11889v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used for scene understanding in autonomous driving, but robustness analysis often relies on task-agnost

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Towards Data-free and Training-free Compression for Speech Foundation Models Using Parameter Clustering

DGX agent

arXiv:2606.11836v1 Announce Type: cross Abstract: This paper presents a novel data-free and training-free compression approach for speech foundation models using channelwise clustering via k-means. Mo

model-releasesarxiv-cs-ai
11 Jun 2026
Research

A Continuous-Time Markov Chain Framework for Insertion Language Models

DGX agent

arXiv:2606.10199v1 Announce Type: cross Abstract: Insertion Language Models (ILMs) offer several advantages over left-to-right generation and mask-based generation. However, existing formulations of i

researcharxiv-cs-cl
10 Jun 2026
Model Releases

Benchmarking stereo reconstruction for 3D printable Martian terrain models

DGX agent

arXiv:2606.10364v1 Announce Type: new Abstract: Reconstructing printable 3D models from Mars rover imagery is challenging because Martian terrain is low-texture, irregular, and partially observed. We

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

CITRAS-FM: Tiny Time Series Foundation Model for Covariate-Informed Zero-Shot Forecasting

DGX agent

arXiv:2606.10798v1 Announce Type: new Abstract: Pretrained time series foundation models (TSFMs) have enabled zero-shot forecasting on unseen target series. However, existing TSFMs often incur high co

model-releasesarxiv-cs-lg
10 Jun 2026
Safety

Conditional Vendi Score: Prompt-Aware Diversity Evaluation for Generative AI Models and LLMs

DGX agent

arXiv:2411.02817v2 Announce Type: replace-cross Abstract: Generative models guided by text prompts are widely evaluated for fidelity and prompt alignment, yet their ability to produce outputs remains

safetyarxiv-cs-ai
10 Jun 2026
Local Ai

Density Field State Space Models: 1-Bit Distillation, Efficient Inference, and Knowledge Organization in Mamba-2

DGX agent

arXiv:2606.10932v1 Announce Type: new Abstract: We present Density Field State Space Models (DF-SSM), a framework for compressing SSMs to a 1-bit scaffold with int8 low-rank correction. Applied to Mam

local-aiarxiv-cs-cl
10 Jun 2026
Safety

Does Reasoning Preserve Alignment? On the Trustworthiness of Large Reasoning Models

DGX agent

arXiv:2606.11046v1 Announce Type: new Abstract: Instruction-tuned LLMs are increasingly converted into reasoning models through post-training to improve multi-step task performance. This conversion is

safetyarxiv-cs-cl
10 Jun 2026
Research

Few-step Generative Models as Lossy Compression

DGX agent

arXiv:2606.10450v1 Announce Type: new Abstract: DiffC provides a principled way to reuse pre-trained diffusion models for lossy compression, but its encoding and decoding procedures remain slow becaus

researcharxiv-cs-cv
10 Jun 2026
Safety

Going with the Flow: Koopman Behavioral Models as Pseudo Planners for Visuo-Motor Dexterity

DGX agent

arXiv:2602.07413v3 Announce Type: replace Abstract: Contemporary visuo-motor dexterity models often rely on expressive policy classes with diffusion and transformer backbones to achieve strong perform

safetyarxiv-cs-ro
10 Jun 2026
Safety

MedFeat: Model-Aware and Explainability-Driven Feature Engineering with LLMs for Clinical Tabular Prediction

DGX agent

arXiv:2603.02221v2 Announce Type: replace-cross Abstract: In clinical tabular prediction, classical machine learning models with feature engineering often outperform neural methods. LLMs are increasin

safetyarxiv-cs-ai
10 Jun 2026
Safety

Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models

DGX agent

arXiv:2606.11167v1 Announce Type: new Abstract: Full-duplex spoken dialogue models can listen and speak simultaneously, making them a promising architecture for natural conversation. However, current

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

OpenRTLSet: A Fully Open-Source Dataset for Large Language Model-based Verilog Module Design

DGX agent

arXiv:2606.10285v1 Announce Type: new Abstract: OpenRTLSet introduces the largest fully open-source dataset for hardware design, offering over 131,000 diverse Verilog code samples to the research comm

model-releasesarxiv-cs-cl
10 Jun 2026
Research

PRISM: Parallel Residual Iterative Sequence Model

DGX agent

arXiv:2602.10796v3 Announce Type: replace Abstract: Generative sequence modeling faces a fundamental tension between the expressivity of Transformers and the efficiency of linear sequence models. Exis

researcharxiv-cs-lg
10 Jun 2026
Model Releases

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models

DGX agent

arXiv:2606.11164v1 Announce Type: new Abstract: Long chain-of-thought (CoT) trajectories in large language model (LLM) reasoning cause severe inference bottlenecks due to rapid key-value (KV) cache gr

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling

DGX agent

arXiv:2606.09926v1 Announce Type: cross Abstract: Sampling from the sequence-level power distribution p^alpha elicits RL-level reasoning from base language models without any parameter updates, but th

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SSR-Merge: Subspace Signal Routing for Training-Free LoRA Merging in Diffusion Models

DGX agent

arXiv:2606.10617v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) merging can efficiently combine diverse generative capabilities from multiple trained LoRAs for a diffusion model. However, e

model-releasesarxiv-cs-cv
10 Jun 2026
Safety

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models

DGX agent

arXiv:2606.10740v1 Announce Type: new Abstract: Failures in multi-turn reasoning models are largely invisible to terminal-score evaluation. A model can lock onto an unsafe stance early in a long dialo

safetyarxiv-cs-ai
10 Jun 2026
Research

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling

DGX agent

arXiv:2512.14614v2 Announce Type: replace Abstract: This paper presents WorldPlay, a streaming video diffusion model that enables real-time, interactive world modeling with long-term geometric consist

researcharxiv-cs-cv
10 Jun 2026
Research

3D Oral Modelling with Improved Vertex Distribution Using Matching-Based Learning

DGX agent

arXiv:2606.07907v1 Announce Type: cross Abstract: In our previous work, a deep learning-based framework for 3D intraoral reconstruction was proposed. The model directly predicts explicit 3D point clou

researcharxiv-cs-ai
9 Jun 2026
Local Ai

A systematic investigation of molecular encoding methods for drug property predictions across neural network and Transformer encoder-based model

DGX agent

arXiv:2606.08973v1 Announce Type: cross Abstract: Fundamental investigations into how different molecular encoding methods affect molecular property prediction remain relatively limited. In this study

local-aiarxiv-cs-lg
9 Jun 2026
Applications

Adversarial Robustness of Activation Steering in Large Language Models

DGX agent

arXiv:2606.07696v1 Announce Type: cross Abstract: Activation steering has become a popular training-free method to control LLM behavior by injecting precomputed direction vectors into the model's resi

applicationsarxiv-cs-ai
9 Jun 2026
Agents

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models

DGX agent

arXiv:2606.08952v1 Announce Type: new Abstract: Multimodal Foundation Models (MFMs) have made substantial progress, yet remain fragile in spatial reasoning over the physical world. A key bottleneck li

agentsarxiv-cs-ai
9 Jun 2026
Research

An Effective Router for Vision-Language Model Selection

DGX agent

arXiv:2606.08970v1 Announce Type: new Abstract: Vision-language models (VLMs) with varying performance and resource requirements are widely deployed, making it difficult for users to select the most a

researcharxiv-cs-ai
9 Jun 2026
Research

ATM: Action-Consistency Transfer Matrix for Diagnosing and Improving Latent World Models

DGX agent

arXiv:2606.09028v1 Announce Type: cross Abstract: Latent world models are increasingly used for control and goal-conditioned planning, yet assessing whether their learned representations are useful fo

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Benchmarking Vision-Language-Action Models on SO-101: Failure and Recovery Analysis

DGX agent

arXiv:2606.08881v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated strong generalization in robotic manipulation, yet existing evaluations are primarily conducted

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Breaking the Tokenizer Barrier: On-Policy Distillation across Model Families

DGX agent

arXiv:2606.09456v1 Announce Type: new Abstract: On-Policy Distillation (OPD) has become a core technique in the post-training of Large Language Models (LLMs) for transferring knowledge from domain exp

safetyarxiv-cs-lg
9 Jun 2026
Safety

Bridging Traditional Explainability Methods and Multimodal Multilingual Models: An XAI-Based Analysis

DGX agent

arXiv:2606.07533v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) effectively integrate text and audio to interpret context in complex interactive dialogues. However, the inte

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Coarse-to-Fine Hierarchical Alignment for UAV-based Human Detection using Diffusion Models

DGX agent

arXiv:2512.13869v3 Announce Type: replace Abstract: Training object detectors demands extensive, task-specific annotations, yet this requirement becomes impractical in UAV-based human detection due to

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models

DGX agent

arXiv:2606.09142v1 Announce Type: cross Abstract: Egocentric vision offers a first-person view of human perception and decision making, yet its potential for traffic-safety prediction remains underexp

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Diverse Thinking Schemata Elicit Better Reasoning in Large Language Models

DGX agent

arXiv:2606.08974v1 Announce Type: new Abstract: Large reasoning models (LRMs) have attracted increasing attention for their ability to solve complex mathematical problems by generating extended reason

safetyarxiv-cs-ai
9 Jun 2026
Research

Do Video Foundation Models Understand Intuitive Physics? A Layerwise Probing Analysis

DGX agent

arXiv:2606.09646v1 Announce Type: cross Abstract: We study whether pretrained video foundation models encode intuitive-physics information in their frozen representations, and how this information var

researcharxiv-cs-ai
9 Jun 2026
Research

DynaCF: Mitigating Shortcut Learning in Reward Models via Dynamic Counterfactual Sensitivity

DGX agent

arXiv:2606.09043v1 Announce Type: new Abstract: Reward models trained from pairwise preferences often exploit superficial shortcut cues rather than learning true response quality. We propose DynaCF, a

researcharxiv-cs-lg
9 Jun 2026
Research

Echo-Memory: A Controlled Study of Memory in Action World Models

DGX agent

arXiv:2606.09803v1 Announce Type: new Abstract: We present extbf{Echo-Memory}, a controlled study of memory mechanisms in action-conditioned world models. These models generate multi-segment videos fr

researcharxiv-cs-cv
9 Jun 2026
Tutorials

Emergence of Context Characteristics Sensitivity in Large Language Models

DGX agent

arXiv:2606.09525v1 Announce Type: cross Abstract: During instruction fine-tuning (IFT), large language models (LLMs) learn to follow instructions by using the provided context to answer a query. While

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

Enhancing Spatial Reasoning in Large Language Models for Metal-Organic Frameworks Structure Prediction

DGX agent

arXiv:2601.09285v2 Announce Type: replace Abstract: Metal-organic frameworks (MOFs) are porous crystalline materials with broad applications such as carbon capture and drug delivery, yet accurately pr

model-releasesarxiv-cs-lg
9 Jun 2026
Research

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling

DGX agent

arXiv:2606.08302v1 Announce Type: new Abstract: Visual Autoregressive (VAR) models adopt a next-scale prediction paradigm, offering high-quality generation with substantially fewer decoding steps. How

researcharxiv-cs-cv
9 Jun 2026
Model Releases

How Much Dense Attention is Necessary? Oracle-Guided Sparse Prefill for Full/GQA Layers in Hybrid Long-Context Models

DGX agent

arXiv:2606.07703v1 Announce Type: cross Abstract: Long-context prefill remains expensive because full/GQA layers still score the historical sequence, even in hybrid models with local, sparse, linear,

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

How Well Do Latent World Models Understand Partially Observable Safety Constraints?

DGX agent

arXiv:2510.06492v2 Announce Type: replace Abstract: Latent world models are a promising approach for learning state representations and dynamics directly from high-dimensional observations, enabling r

safetyarxiv-cs-ro
9 Jun 2026
← Previous
1…104105106107108…1030
Next →