AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

DiffAttack: Evasion Attacks Against Face Recognition via Latent Diffusion Models

DGX agent

arXiv:2607.28936v1 Announce Type: cross Abstract: Facial biometric identification relies on the distinctiveness of user attributes within a high-dimensional embedding space. However, the decision boun

model-releasesarxiv-cs-ai
3 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Fisher Information, Training and Bias in Fourier Regression Models

DGX agent

arXiv:2510.06945v2 Announce Type: replace Abstract: Motivated by the growing interest in quantum machine learning, in particular quantum neural networks (QNNs), we study how recently introduced evalua

safetyarxiv-cs-lg
3 Aug 2026
Model Releases

RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model

DGX agent

arXiv:2607.17977v2 Announce Type: replace Abstract: We present RynnBrain 1.1, a family of embodied foundation models spanning 2B, 9B, and 122B-A10B scales. Trained with a unified spatio-temporal and p

model-releasesarxiv-cs-ro
3 Aug 2026
Research

When Model Priors Conflict with Visual Evidence: Mitigating Commonsense-Driven Hallucinations by Selective Prior Calibration

DGX agent

arXiv:2607.29240v1 Announce Type: cross Abstract: In vision--language models, commonsense-driven hallucination (CDH) occurs when a model's commonsense prior overrides clear visual evidence of an atypi

researcharxiv-cs-ai
3 Aug 2026
Safety

Ask don't tell: Reducing sycophancy in large language models

DGX agent

arXiv:2602.23971v4 Announce Type: replace-cross Abstract: Sycophancy, the tendency of large language models to favour user-affirming responses over critical engagement, has been identified as an align

safetyarxiv-cs-ai
31 Jul 2026
Model Releases

Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering

DGX agent

arXiv:2607.28568v1 Announce Type: new Abstract: Recursive self-improvement (RSI) requires AI systems that improve the process of building AI (i.e., AI4AI); machine learning engineering (MLE) offers a

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

KAISEN: Reproducible Subgroup Fairness Auditing for Clinical Risk Models

DGX agent

arXiv:2607.28608v1 Announce Type: new Abstract: Clinical risk models routinely achieve strong aggregate performance while producing materially different error rates across patient subgroups. Audit pip

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

MedHallTune: An Instruction-Tuning Benchmark for Mitigating Medical Hallucination in Vision-Language Models

DGX agent

arXiv:2502.20780v2 Announce Type: replace-cross Abstract: The increasing use of vision-language models (VLMs) in healthcare applications presents great challenges related to hallucinations, in which t

model-releasesarxiv-cs-cl
31 Jul 2026
Research

MRD: Using Physically Based Differentiable Rendering to Probe Vision Models for 3D Scene Understanding

DGX agent

arXiv:2512.12307v5 Announce Type: replace Abstract: While deep learning methods have achieved impressive success in many vision benchmarks, it remains difficult to understand and explain the represent

researcharxiv-cs-cv
31 Jul 2026
Model Releases

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models

DGX agent

arXiv:2607.28609v1 Announce Type: cross Abstract: Computer-using agents (CUAs) are advancing rapidly across the digital world. A CUA trajectory records the agent's actions, states, and reasoning. Veri

model-releasesarxiv-cs-cl
31 Jul 2026
Research

QuantWAMs: Calibrating at the Right Granularity for World Action Models

DGX agent

arXiv:2607.28405v1 Announce Type: cross Abstract: World Action Models (WAMs) jointly predict future observations and actions, but their iterative denoising and closed-loop execution make efficient dep

researcharxiv-cs-lg
31 Jul 2026
Model Releases

Scaling Vision-Language Models Is Not Enough to Mitigate Bias

DGX agent

arXiv:2607.28211v1 Announce Type: new Abstract: Vision-Language Models (VLMs) such as CLIP are now foundational to multimodal systems, yet their robustness to spurious correlations remains poorly unde

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Subtract or Replay? Exact Deletion from Language-Model Memory

DGX agent

arXiv:2607.27539v1 Announce Type: cross Abstract: Exact deletion from persistent language-model memory depends on how that memory represents a record. Addressable influence can be removed by algebraic

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

The Confidence Manifold: Geometric Structure of Correctness Representations in Language Models

DGX agent

arXiv:2602.08159v2 Announce Type: replace-cross Abstract: When a language model asserts that 'the capital of Australia is Sydney,' does it know this is wrong? Models assert misconceptions with the sam

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

TriShield: Zero-Utility-Loss Defense Against Privacy Backdoors in Federated Language Model Fine-Tuning via Orthogonal Gradient Projection and Optimizer State Entanglement

DGX agent

arXiv:2607.27940v1 Announce Type: cross Abstract: Federated fine-tuning of large language models (LLMs) enables collaborative training without exposing raw data. However, a recent attack, NeuroImprint

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Existence-Field Diffusion Model for Spatial Point Processes with Variable Cardinality

DGX agent

arXiv:2607.26428v1 Announce Type: new Abstract: We study generative modeling of spatial point processes (SPP), where both the number of points and their spatial configuration are governed by a joint d

researcharxiv-cs-lg
30 Jul 2026
Model Releases

Foundation Models for Face Presentation Attack Detection: A Unified Linear-Probing Benchmark

DGX agent

arXiv:2607.26993v1 Announce Type: new Abstract: Face presentation attack detection (PAD) remains challenging under cross-dataset evaluation, where domain shift degrades models trained on a single data

model-releasesarxiv-cs-lg
30 Jul 2026
Safety

Learning Implicit Causal World Models from Multi-Agent Demonstrations

DGX agent

arXiv:2607.26336v1 Announce Type: new Abstract: In model-based reinforcement learning, world models exist as internal simulators, but their training often conflates statistical correlations with causa

safetyarxiv-cs-lg
30 Jul 2026
Safety

Probing the Origins of Reasoning Performance: Representational Quality for Mathematical Problem-Solving in RL vs. SFT Fine-Tuned Models

DGX agent

arXiv:2607.26119v1 Announce Type: cross Abstract: Large reasoning models trained via reinforcement learning (RL) have been increasingly shown to outperform their supervised fine-tuned (SFT) counterpar

safetyarxiv-cs-cl
30 Jul 2026
Research

VideoNorms: Benchmarking Cultural Awareness of Video Language Models

DGX agent

arXiv:2510.08543v2 Announce Type: replace-cross Abstract: As Video Large Language Models (VideoLLMs) are deployed globally, it is important to assess their ability to reason across cultural contexts.

researcharxiv-cs-cl
30 Jul 2026
Research

Walk through Paintings: Egocentric World Models from Internet Priors

DGX agent

arXiv:2601.15284v2 Announce Type: replace Abstract: What if a video generation model could not only imagine a plausible future, but the correct one -- accurately reflecting how the world changes with

researcharxiv-cs-cv
30 Jul 2026
Safety

Architectural Backdoors in Vision-Language Model Supply Chains via Representation Steering

DGX agent

arXiv:2607.25479v1 Announce Type: cross Abstract: Vision--Language Models (VLMs) are increasingly deployed through a model supply chain in which pretrained checkpoints, architecture definitions, text

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

Building Large-Scale English-Romanian Literary Translation Resources with Open Models

DGX agent

arXiv:2509.07829v4 Announce Type: replace-cross Abstract: Literary translation has recently gained attention as a distinct and complex task in machine translation research, yet translation by small op

model-releasesarxiv-cs-ai
29 Jul 2026
Local Ai

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models

DGX agent

arXiv:2607.25522v1 Announce Type: cross Abstract: The rapid advancement of video generation models has led to the increasing misuse of image-to-video (I2V) models. Although substantial progress has be

local-aiarxiv-cs-ai
29 Jul 2026
Model Releases

Influence of Prompt Engineering on Small Language Models for Guarded Query Routing

DGX agent

arXiv:2607.24801v1 Announce Type: cross Abstract: We study the problem of guarded query routing, where we assume that a user query first meets a router that either determines the ideal endpoint for in

model-releasesarxiv-cs-cl
29 Jul 2026
Research

Measuring and Improving Behavioral Consistency in Large Language Models through Fact-Heuristic-Emotion State Enforcement

DGX agent

arXiv:2607.24765v1 Announce Type: cross Abstract: Large language models (LLMs) can give different answers to the same decision problem across runs, and reverse a decision when their own prior answer r

researcharxiv-cs-ai
29 Jul 2026
Research

MODUS: Decoder-Only Any-to-Any Modeling of Diverse Modalities

DGX agent

arXiv:2607.25948v1 Announce Type: cross Abstract: Any-to-any models predict any modality from any combination of others within a single network, a formulation used in multimodal vision and vision-lang

researcharxiv-cs-ai
29 Jul 2026
Local Ai

TaylorPODA: A Taylor Expansion-Based Method to Improve Post-Hoc Attributions for Opaque Models

DGX agent

arXiv:2507.10643v4 Announce Type: replace-cross Abstract: Post-hoc model-agnostic local attribution (LA) methods have been widely adopted to explain opaque AI models by quantifying feature-wise contri

local-aiarxiv-cs-ai
29 Jul 2026
Model Releases

AgentOmnia: Scaling Agentic Models for Full-Scenario Applications

DGX agent

arXiv:2607.23124v1 Announce Type: new Abstract: Large language model agents have advanced rapidly, yet progress remains fragmented across domains, capabilities, task difficulty, and interaction settin

model-releasesarxiv-cs-ai
28 Jul 2026
Research

AIFL: A Global Daily Streamflow Forecasting Model Using a Deterministic LSTM Pre-trained on ERA5-Land and Fine-tuned on IFS

DGX agent

arXiv:2602.16579v2 Announce Type: replace-cross Abstract: Reliable global streamflow forecasting is essential for flood preparedness and water resource management, yet data-driven models often suffer

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Algorithmic Blindness in Large Language Models: A Calibration Study of Performance Prediction

DGX agent

arXiv:2602.21947v5 Announce Type: replace Abstract: Large language models (LLMs) demonstrate remarkable breadth of knowledge, yet their ability to reason about computational processes remains poorly u

model-releasesarxiv-cs-cl
28 Jul 2026
Agents

AutoWorld: Learning Multi-Agent Traffic Simulation with Self-Supervised World Models

DGX agent

arXiv:2603.28963v2 Announce Type: replace-cross Abstract: Simulation with realistic traffic agents is essential for validating autonomous driving systems. Existing data-driven simulators learn agent b

agentsarxiv-cs-ai
28 Jul 2026
Research

DomainPilot: Domain-Level Loss-Guided Two-Stage Data Mixture Optimization for Efficient Language Model Fine-Tuning

DGX agent

arXiv:2607.22769v1 Announce Type: cross Abstract: The training efficacy of large language models (LLMs) is fundamentally constrained by the quality and composition of training data. Existing dynamic d

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Embodied GPT-5.1: Evidence of a World Model?

DGX agent

arXiv:2607.23899v1 Announce Type: cross Abstract: This exploratory study examines whether a large multimodal language model, GPT-5.1, can serve as the high-level controller of a physical mobile robot

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Extracting Algorithms in Pre-trained LLMs: A Case on Hidden Markov Models

DGX agent

arXiv:2607.22646v1 Announce Type: new Abstract: Large language models (LLMs) display a striking ability to predict next observations from Hidden Markov Models (HMMs) via in-context learning (ICL), but

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Gaze-Anchored Social Net: Decoding Implicit Relations via Joint Modeling

DGX agent

arXiv:2607.22847v1 Announce Type: new Abstract: Human gaze does more than point to visual targets; it serves as a subtle indicator of social intent within static images, whereas standard models typica

model-releasesarxiv-cs-cv
28 Jul 2026
Research

Generative Artificial Intelligence (GenAI) to convert images of queuing networks into verifiable simulation models: an open-weight LLM workflow approach

DGX agent

arXiv:2607.24259v1 Announce Type: new Abstract: Recent work has explored the use of Large Language Models (LLMs) to automate simulation model building, typically by generating executable code directly

researcharxiv-cs-ai
28 Jul 2026
Model Releases

GOTS: Greedy Orthogonal Token Selection for High-Resolution Vision-Language Models

DGX agent

arXiv:2607.23913v1 Announce Type: new Abstract: Modern vision-language models (VLMs) increasingly rely on dynamic or high-resolution visual encoding, producing thousands of visual tokens that substant

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Gubernaut: A Deterministic Homeostatic Controller for Affect-Regulated LLM Agents, Validated Across Independent Model Families

DGX agent

arXiv:2607.24339v1 Announce Type: new Abstract: Large language model (LLM) agents inherit reactive failure modes: escalation under provocation, sycophantic drift under flattery, perseveration when stu

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Infinite-Precision Autoregressive Modeling for Vector Graphics and Layouts

DGX agent

arXiv:2601.05680v2 Announce Type: replace-cross Abstract: While Transformer-based autoregressive models excel in data generation, their token discretization strategy inherently limits their precision

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models

DGX agent

arXiv:2607.24273v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong potential in financial reasoning, but existing benchmarks often evaluate domain knowledge, numerical reas

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Multi-model approach for autonomous driving: A comprehensive study on traffic sign-, vehicle- and lane detection and behavioral cloning

DGX agent

arXiv:2603.09255v2 Announce Type: replace-cross Abstract: Deep learning and computer vision techniques have become increasingly important in the development of self-driving cars. These techniques play

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Reconstructing Item Characteristic Curves using Fine-Tuned Large Language Models

DGX agent

arXiv:2601.02580v2 Announce Type: replace-cross Abstract: Traditional methods for determining assessment item parameters, such as difficulty and discrimination, rely heavily on expensive field testing

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

SketchMamba: A Lightweight State-Space Model for Joint Progressive Sketch Classification and Stroke Auto-Completion

DGX agent

arXiv:2607.23580v1 Announce Type: new Abstract: Existing vector-sketch models treat recognition and generation as separate tasks, leaving a gap for streaming interfaces that must understand a drawing

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Sparse Gaussian-Mixture-Model Q-Functions via Hadamard Overparametrization for Online Reinforcement Learning

DGX agent

arXiv:2607.23474v1 Announce Type: new Abstract: This paper develops an online, off-policy policy-iteration framework for reinforcement learning (RL), based on sparse Gaussian-mixture-model Q-functions

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

Who Gets Named: Citation Type Predicts Individual Naming by Grounded Language Models, and a Roster Instrument Captures 0.5% of It

DGX agent

arXiv:2607.23893v1 Announce Type: cross Abstract: Prior work on AI brand visibility measures the firm: does a model recommend a company, and does that track its reputation. This study asks the questio

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study

DGX agent

arXiv:2607.21988v1 Announce Type: new Abstract: Self-harm content is particularly challenging to detect using NLP techniques, and is also a high-stakes task which requires the highest accuracy to enab

model-releasesarxiv-cs-cl
27 Jul 2026
Safety

Autoregressive EHR Foundation Models with Multimodal Inputs

DGX agent

arXiv:2607.22264v1 Announce Type: new Abstract: Autoregressive foundation models trained on tokenized electronic health records (EHRs) can support zero-shot clinical prediction, yet most operate on st

safetyarxiv-cs-lg
27 Jul 2026
← Previous
1…3132333435…1021
Next →