AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

FPEdit: Robust LLM Fingerprinting through Localized Parameter Editing

DGX agent

arXiv:2508.02092v3 Announce Type: replace-cross Abstract: Large language models represent significant investments in computation, data, and engineering expertise, making them extraordinarily valuable

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PRISM Edit: One Vector for All Temporal Answers

DGX agent

arXiv:2607.11327v2 Announce Type: replace-cross Abstract: Model editing keeps large language models (LLMs) up to date without retraining, but temporal facts expose a limitation of the prevailing locat

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Off-the-Shelf LLMs as Process Scorers: Training-Free Alternative to PRMs for Mathematical Reasoning

DGX agent

arXiv:2606.01682v1 Announce Type: cross Abstract: Selecting the best response from multiple small-model samples using a stronger scorer is a simple inference-time strategy, but fails when the small mo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation

DGX agent

arXiv:2606.01481v1 Announce Type: new Abstract: With the rapid advancements in text-to-image diffusion models, generative video models (T2V models) like Sora can now produce short synthetic videos fro

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

jina-embeddings-v5-omni: Text-Geometry-Preserving Multimodal Embeddings via Frozen-Tower Composition

DGX agent

arXiv:2605.08384v1 Announce Type: new Abstract: In this work, we introduce frozen-encoder model composition, a novel approach to multimodal embedding models. We build on the VLM-style architecture, in

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

A Probabilistic Consensus-Driven Approach for Robust Counterfactual Explanations

DGX agent

arXiv:2604.17494v1 Announce Type: new Abstract: Counterfactual explanations (CFEs) are essential for interpreting black-box models, yet they often become invalid when models are slightly changed. Exis

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting

DGX agent

arXiv:2604.18452v1 Announce Type: cross Abstract: Vision-language modeling is rapidly increasing in popularity with an ever expanding list of available models. In most cases, these vision-language mod

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

User-Assistant Bias in LLMs

DGX agent

arXiv:2508.15815v3 Announce Type: replace Abstract: Modern large language models (LLMs) are typically trained and deployed using structured role tags (e.g. system, user, assistant, tool) that explicit

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

SLM Finetuning for Natural Language to Domain Specific Code Generation in Production

DGX agent

arXiv:2604.09952v1 Announce Type: new Abstract: Many applications today use large language models for code generation; however, production systems have strict latency requirements that can be difficul

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction

DGX agent

arXiv:2407.08101v4 Announce Type: replace Abstract: Vision-language models have shown impressive progress in recent years. However, existing models are largely limited to turn-based interactions, wher

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

EXAONE 4.5 Technical Report

DGX agent

arXiv:2604.08644v1 Announce Type: new Abstract: This technical report introduces EXAONE 4.5, the first open-weight vision language model released by LG AI Research. EXAONE 4.5 is architected by integr

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

A Bayes-Markov Neuromorphic Model of Cortical Orientation Selectivity: A Computational Re-implementation and Quantitative Simulation Study

DGX agent

arXiv:2608.12388v1 Announce Type: cross Abstract: The emergence of orientation selectivity in the primary visual cortex (V1) remains a central question in computational neuroscience. Shirazi's Bayes-M

model-releasesarxiv-cs-lg
14 Aug 2026
Model Releases

Are Large Language Models Reliable Reviewers? A Benchmark for Error Detection in Financial Documents

DGX agent

arXiv:2608.12342v1 Announce Type: new Abstract: Ensuring the accuracy of financial documents is critical for economic analysis, regulatory compliance, and corporate decision-making. Several studies ha

model-releasesarxiv-cs-cl
14 Aug 2026
Research

Are you Talking Logic to Me? Assessing Language Models Syllogistic Reasoning Capabilities

DGX agent

arXiv:2608.12374v1 Announce Type: cross Abstract: Language models (LMs) struggle with logical tasks like reasoning on syllogisms. It has been shown that Knowledge Representation (KR) plays a crucial r

researcharxiv-cs-ai
14 Aug 2026
Model Releases

Behavioral Reprogramming of Open-Weights Models: Cognitive Plasticity and Alignment Bounds

DGX agent

arXiv:2608.13069v1 Announce Type: new Abstract: Large language models (LLMs) are predominantly aligned to function as passive, sycophantic assistants. We challenge this default paradigm by empirically

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

FIRE-VLA: Failure-Informed Self-Evolution for Vision-Language-Action Models in Autonomous Driving

DGX agent

arXiv:2608.13395v1 Announce Type: new Abstract: Reinforcement learning improves autonomous-driving vision-language-action (VLA) models by evaluating trajectories sampled from the current policy. Group

model-releasesarxiv-cs-ro
14 Aug 2026
Tutorials

Jointly Predicting Courses and Grades Using a Transformer-Based Model

DGX agent

arXiv:2608.13409v1 Announce Type: new Abstract: Existing predictive models in learning analytics often treat student academic history as a simple sequence, overlooking the concurrent nature of courses

tutorialsarxiv-cs-ai
14 Aug 2026
Model Releases

Large Language Models Pass the History Exam But Miss the <<History>>: A Polish High School Exit Exam Matura Benchmark

DGX agent

arXiv:2608.12343v1 Announce Type: new Abstract: AI chatbots are widely used by students as knowledge sources, yet LLM benchmarks rarely assess interpretative historical reasoning. We evaluate eight le

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

Novels generated by language models show compressed formal variation

DGX agent

arXiv:2608.12630v1 Announce Type: cross Abstract: While large language models can generate entire novels, there is little information about the level of formal variation in their output over many gene

model-releasesarxiv-cs-ai
14 Aug 2026
Research

Personalized Scorer Modeling: A Learning-Based Framework for Deriving Robust Sleep Stage Labels from Multiple Experts

DGX agent

arXiv:2608.12446v1 Announce Type: cross Abstract: Sleep stage classification is important for the diagnosis and management of sleep disorders, yet most automatic staging studies evaluate models agains

researcharxiv-cs-ai
14 Aug 2026
Safety

Scaling Automatic Research Agents via World Models

DGX agent

arXiv:2608.12564v1 Announce Type: new Abstract: Automating empirical research is a long-standing direction of AI. Recent automatic research (AutoResearch) agents bring this goal within reach, as moder

safetyarxiv-cs-lg
14 Aug 2026
Model Releases

Air Quality Station Simulation via LSTM and Attention-Based Modelling

DGX agent

arXiv:2608.11839v1 Announce Type: new Abstract: Poor air quality in urban areas is driven by a complex chain of processes and presents a significant public health concern. To better understand and con

model-releasesarxiv-cs-lg
13 Aug 2026
Research

Can Vision Models Read the Radar Display? On the Feasibility of Radar Imagery for Air Traffic Complexity Estimation

DGX agent

arXiv:2608.11810v1 Announce Type: new Abstract: Air traffic controllers perceive traffic complexity through the radar display, suggesting that a computer vision model operating on the same imagery may

researcharxiv-cs-cv
13 Aug 2026
Model Releases

CT-DeltaBench: A Benchmark for Longitudinal 3D Medical Imaging Difference Reporting with Vision-Language Models

DGX agent

arXiv:2608.11534v1 Announce Type: new Abstract: In medical imaging, the clinical value of Computed Tomography (CT) lies not only in depicting current disease status, but crucially in enabling longitud

model-releasesarxiv-cs-cl
13 Aug 2026
Research

How effective are VLMs in assisting humans in inferring the quality of mental models from Multimodal short answers?

DGX agent

arXiv:2603.00056v2 Announce Type: replace-cross Abstract: STEM Mental models can play a critical role in assessing students' conceptual understanding of a topic. They not only offer insights into what

researcharxiv-cs-ai
13 Aug 2026
Model Releases

Orientation, not magnitude: the causal structure of task-vector interference in merged language models

DGX agent

arXiv:2608.11797v1 Announce Type: new Abstract: Model merging by task arithmetic works until it doesn't, and the field diagnoses why with magnitudes: layerwise representation bias, deviations from cro

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop

DGX agent

arXiv:2608.11215v1 Announce Type: new Abstract: Simulating societies of many large language model (LLM) agents is expensive, yet the questions asked of such simulations are usually macroscopic: phase

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Qwen-MusicAVQA-7B: A Multimodal Model for Music Audio-Visual QA

DGX agent

arXiv:2608.11329v1 Announce Type: cross Abstract: A common approach to adding audio to a vision-language model is to train or adapt a large omni-modal system. We show that a lightweight alternative ca

model-releasesarxiv-cs-cv
13 Aug 2026
Model Releases

Semantic Lenia: Emergence of Homeostatic Solitons within the Semantic Space of Large Language Models

DGX agent

arXiv:2608.11657v1 Announce Type: cross Abstract: We introduce Semantic Lenia, an artificial life framework that transforms Large Language Model (LLM) inference from a static optimization problem into

model-releasesarxiv-cs-ai
13 Aug 2026
Local Ai

Certify or Refuse: A Cross-Model Map for Selective Risk Control with Coverage Floors under Covariate Shift

DGX agent

arXiv:2608.10893v1 Announce Type: new Abstract: Certified selective predictors attain whatever coverage they attain; operators impose an automation floor: answer at least a eta-fraction of shifted tar

local-aiarxiv-cs-cl
12 Aug 2026
Safety

Hidden in Plain Sight: Diffusion-Based Unrestricted Robotic Attacks on Vision-Language-Action Models

DGX agent

arXiv:2608.10393v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong capabilities in controlling robots across diverse manipulation tasks. However, their adversarial r

safetyarxiv-cs-ai
12 Aug 2026
Research

HQ-DM: Single Hadamard Transformation-Based Quantization-Aware Training for Low-Bit Diffusion Models

DGX agent

arXiv:2512.05746v3 Announce Type: replace Abstract: Diffusion models have demonstrated significant applications in the field of image generation. However, their high computational and memory costs pos

researcharxiv-cs-cv
12 Aug 2026
Model Releases

Locally Deployable Small Language Models for Emergency Department Decision Support: A Systematic Benchmark of Fine-Tuning Strategies

DGX agent

arXiv:2608.10273v1 Announce Type: cross Abstract: Deploying large language models (LLMs) for decision support in emergency departments (EDs) faces two major challenges: privacy risks of transmitting p

model-releasesarxiv-cs-ai
12 Aug 2026
Tutorials

ReCBM: Uncertainty-Gated Relational Reasoning for Concept Bottleneck Models

DGX agent

arXiv:2608.10004v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) provide an interpretable framework by grounding predictions in human-understandable concepts, enabling semantic inspect

tutorialsarxiv-cs-ai
12 Aug 2026
Safety

Surgical WAM: A World-Action Model for Data-Efficient Surgical Robot Learning

DGX agent

arXiv:2608.11204v1 Announce Type: cross Abstract: Learning reliable surgical manipulation policies is bottlenecked by the scarcity of action-labeled demonstrations: teleoperated surgical robot (e.g.,

safetyarxiv-cs-ai
12 Aug 2026
Agents

Activation Probes Surface Code-Security Signals that the Model's Output Misses

DGX agent

arXiv:2608.09643v1 Announce Type: cross Abstract: AI coding agents now write a growing share of production code, and human security review does not scale at the rate code is generated. The agents in w

agentsarxiv-cs-lg
11 Aug 2026
Model Releases

An Agentic AI Framework Overcomes Fundamental Limitations of Large Language Models for Glaucoma Detection from Fundus Photography

DGX agent

arXiv:2608.07651v1 Announce Type: new Abstract: Large language models (LLMs) show promise in medical image interpretation but suffer from hallucination, limited accuracy, and run-to-run inconsistency.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CosmosAlign: Adapting a World Foundation Model for Generative Traffic Video Forecasting

DGX agent

arXiv:2608.07693v1 Announce Type: cross Abstract: Generative traffic video forecasting aims to synthesize long-horizon, temporally coherent future videos of traffic scenes from a short observation his

model-releasesarxiv-cs-ai
11 Aug 2026
Research

DialectS2S: End-to-End Speech Dialogue Modeling for Low-Resource Chinese Dialects

DGX agent

arXiv:2608.08067v1 Announce Type: cross Abstract: Current end-to-end speech dialogue models are primarily optimized for mainstream languages and remain limited in low-resource dialect scenarios due to

researcharxiv-cs-ai
11 Aug 2026
Tutorials

Ensemble learning of pathology foundation models for precision oncology

DGX agent

arXiv:2508.16085v2 Announce Type: replace Abstract: Histopathology is essential for cancer diagnosis and treatment selection, and pathology foundation models learn visual representations from whole-sl

tutorialsarxiv-cs-cv
11 Aug 2026
Model Releases

FemWear: A Specialized Wearable Foundation Model for Women's Health

DGX agent

arXiv:2608.08244v1 Announce Type: new Abstract: General wearable foundation models are pretrained across broad sensor streams and populations, but are not designed around women's-health tasks. We intr

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

FitAQA: A Benchmark of Fitness Action Quality Assessment for Multimodal Large Language Models

DGX agent

arXiv:2608.08736v1 Announce Type: new Abstract: Fitness Action Quality Assessment (AQA) is important for intelligent sports training, yet the capabilities of Multimodal Large Language Models (MLLMs) i

model-releasesarxiv-cs-ai
11 Aug 2026
Research

From Independent to Correlated Diffusion: Generalized Generative Modeling with Probabilistic Computers

DGX agent

arXiv:2603.27996v2 Announce Type: replace Abstract: Diffusion models have emerged as a powerful framework for generative tasks in deep learning. They decompose generative modeling into two computation

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Fusion Training for Mathematical Generalization in Large Language Models

DGX agent

arXiv:2608.09893v1 Announce Type: cross Abstract: Thinking Mode Fusion (TMF) enables large language models to support both concise responses and long-form reasoning by unifying a non-thinking mode and

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

GWM-VLA: Geometry-Aware Latent World Modeling for Vision-Language-Action Learning

DGX agent

arXiv:2608.07619v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models achieve strong robotic manipulation performance but often degrade under visual and environmental shifts. Latent worl

local-aiarxiv-cs-ro
11 Aug 2026
Research

MoE-Prism: Disentangling Monolithic Experts for Elastic MoE Services via Model-System Co-Designs

DGX agent

arXiv:2510.19366v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) scales model capacity through sparse activation, and is becoming an important architecture for large language models (LLMs)

researcharxiv-cs-cl
11 Aug 2026
Model Releases

MonitorBench: A Comprehensive Benchmark for Chain-of-Thought Monitorability in Large Language Models

DGX agent

arXiv:2603.28590v3 Announce Type: replace Abstract: Large language models (LLMs) can generate chains of thought (CoTs) that are not always causally responsible for their final outputs. When such a mis

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Omni2LoRA: Coherence-Preserving Parametric Memory for Efficient Omni Language Models

DGX agent

arXiv:2608.09227v1 Announce Type: new Abstract: Omnimodal language models (OLMs) enable unified audio-visual understanding, but processing long joint token sequences makes inference computationally pr

model-releasesarxiv-cs-ai
11 Aug 2026
← Previous
1…7576777879…1030
Next →