AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

Diagnosing Visual Ignorance in Vision-Language Models

DGX agent

arXiv:2606.06890v1 Announce Type: new Abstract: Vision-Language Models (VLMs) frequently rely on language priors, producing confident answers that are weakly grounded in visual evidence. While this be

model-releasesarxiv-cs-cv
8 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Elmes*: Automated Construction of Fine-Grained Evaluation Rubrics for Large Language Models in Long-Tail Educational Scenarios

DGX agent

arXiv:2606.06546v1 Announce Type: new Abstract: Evaluating large language models (LLMs) for education requires measuring how models teach, not only what they know. Existing benchmarks emphasize domain

safetyarxiv-cs-lg
8 Jun 2026
Model Releases

GuideCAD: A Lightweight Multimodal Framework for 3D CAD Model Generation via Prefix Embedding

DGX agent

arXiv:2606.07024v1 Announce Type: new Abstract: Multi-modal approaches used for 3D CAD generation require substantial computational resources, necessitating efficient training. To address this, we pro

model-releasesarxiv-cs-cv
8 Jun 2026
Research

Finite Element-Based Material Learning via Automatic Differentiation: Learning constitutive neural network models from full-field deformation data

DGX agent

arXiv:2606.05199v1 Announce Type: cross Abstract: The identification of constitutive neural network models from heterogeneous full-field deformation data provides a robust alternative to traditional c

researcharxiv-cs-ai
6 Jun 2026
Safety

LatentWave: JEPA Pretraining for Wireless Foundation Models

DGX agent

arXiv:2606.06373v1 Announce Type: cross Abstract: Wireless foundation models have emerged as a promising alternative to building separate models for each wireless task. However, existing approaches re

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

Minimizing the Hidden Cost of Scales: Graph-Guided Ultra-Low-Bit Quantization for Large Language Models

DGX agent

arXiv:2606.05429v1 Announce Type: new Abstract: Post-training quantization (PTQ) is critical for the efficient deployment of large language models (LLMs). Recent ultra-low-bit PTQ methods rely on rigi

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Towards Unified and Data-Efficient Prognostics and Health Management with Tabular Foundation Models

DGX agent

arXiv:2606.05481v1 Announce Type: cross Abstract: Data-driven Prognostics and Health Management (PHM) uses time-varying condition-monitoring data to diagnose system states and estimate remaining usefu

researcharxiv-cs-ai
6 Jun 2026
Local Ai

AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding

DGX agent

arXiv:2606.06155v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models leverage the rich world knowledge of pretrained vision-language models (VLMs) to enable instruction-following robo

local-aiarxiv-cs-cv
5 Jun 2026
Model Releases

Almieyar-Oryx-BloomBench: A Bilingual Multimodal Benchmark for Cognitively Informed Evaluation of Vision-Language Models

DGX agent

arXiv:2606.05531v1 Announce Type: cross Abstract: Despite the rapid progress of Vision-Language Models (VLMs), the field lacks benchmarks that rigorously diagnose their true reasoning abilities and ch

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Code2LoRA: Hypernetwork-Generated Adapters for Code Language Models under Software Evolution

DGX agent

arXiv:2606.06492v1 Announce Type: cross Abstract: Code language models need repository-level context to resolve imports, APIs, and project conventions. Existing methods inject this knowledge as long i

model-releasesarxiv-cs-cl
5 Jun 2026
Research

MASF: A Multi-Model Adaptive Selection Framework for Abstractive Text summarization

DGX agent

arXiv:2606.05494v1 Announce Type: new Abstract: Automatic text summarization has become increasingly important due to the rapid growth of digital textual information. This paper presents a Multi-Model

researcharxiv-cs-cl
5 Jun 2026
Agents

Merging model-based control with multi-agent reinforcement learning for multi-agent cooperative teaming strategies

DGX agent

arXiv:2606.06011v1 Announce Type: new Abstract: In this work, we propose a framework that combines multi-agent reinforcement learning (MARL) with model-based control to achieve safe, dynamically feasi

agentsarxiv-cs-ro
5 Jun 2026
Model Releases

PlanBench-V: A Spatial Planning Map Benchmark for Vision-Language Models

DGX agent

arXiv:2606.05744v1 Announce Type: new Abstract: Spatial planning maps are central to territorial governance, translating planning objectives, regulations, and spatial strategies into visual forms for

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Seeing is Believing? Evaluating Vision-Language Model Susceptibility in Agent-to-Agent Multimodal Persuasion

DGX agent

arXiv:2510.22768v2 Announce Type: replace Abstract: As autonomous agents increasingly interact, they inevitably attempt to influence one another. While prior work in text-only settings has explored th

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

AutoLab: Can Frontier Models Solve Long-Horizon Auto Research and Engineering Tasks?

DGX agent

arXiv:2606.05080v1 Announce Type: new Abstract: Scientific and engineering progress is fundamentally a long-horizon iterative process: proposing changes, running experiments, measuring outcomes, and c

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Beyond Objective Equivalence: Constraint Injection for LLM-Based Optimization Modeling on Vehicle Routing Problems

DGX agent

arXiv:2606.04816v1 Announce Type: new Abstract: Large language models (LLMs) increasingly translate natural-language optimization problems into executable solver code. Yet for constraint-dense operati

model-releasesarxiv-cs-ai
4 Jun 2026
Research

Data Attribution in Large Language Models via Bidirectional Gradient Optimization

DGX agent

arXiv:2606.04928v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed across diverse applications, raising critical questions for governance, accountability, and dat

researcharxiv-cs-cl
4 Jun 2026
Tutorials

Effective vocabulary expansion of multilingual language models for extremely low-resource languages

DGX agent

arXiv:2602.09388v2 Announce Type: replace Abstract: Multilingual pre-trained language models(mPLMs) offer significant benefits for many low-resource languages. To further expand the range of languages

tutorialsarxiv-cs-cl
4 Jun 2026
Model Releases

Gradient estimators for parameter inference in discrete stochastic kinetic models

DGX agent

arXiv:2604.02121v2 Announce Type: replace-cross Abstract: Stochastic kinetic models are ubiquitous in physics, yet inferring their parameters from experimental data remains challenging. For determinis

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Hierarchical Self-Supervised Adversarial Training for Robust Vision Models in Histopathology

DGX agent

arXiv:2503.10629v2 Announce Type: replace Abstract: Adversarial attacks pose significant challenges for vision models in critical fields like healthcare, where reliability is essential. Although adver

model-releasesarxiv-cs-cv
4 Jun 2026
Tutorials

Measuring What Matters: Synthetic Benchmarks for Concept Bottleneck Models

DGX agent

arXiv:2606.04326v1 Announce Type: cross Abstract: Concept bottleneck models predict outcomes from high-level concepts detected in inputs. Although concepts provide a simple way to reap benefits from i

tutorialsarxiv-cs-ai
4 Jun 2026
Safety

OSCAR: Omni-Embodiment Skeleton-Conditioned World Action Model for Robotics

DGX agent

arXiv:2606.04463v1 Announce Type: new Abstract: We present OSCAR, a precise action-conditioned video world model that generalizes across different robot embodiments and enables robot policy evaluation

safetyarxiv-cs-ro
4 Jun 2026
Research

Overclocking Electrostatic Generative Models

DGX agent

arXiv:2509.22454v2 Announce Type: replace Abstract: Electrostatic generative models such as PFGM++ have recently emerged as a powerful framework, achieving competitive performance in image synthesis.

researcharxiv-cs-lg
4 Jun 2026
Safety

POLARIS: Guiding Small Models to Write Long Stories

DGX agent

arXiv:2606.04095v1 Announce Type: cross Abstract: Small open-weight models struggle at long-form creative writing: their generated stories either fall far short of the requested length, or their quali

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay

DGX agent

arXiv:2603.23841v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) are increasingly used as primary sources of information, their potential for political bias may impact thei

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Prompt-Level Distillation: A Non-Parametric Alternative to Model Fine-Tuning for Efficient Reasoning

DGX agent

arXiv:2602.21103v2 Announce Type: replace Abstract: Advanced reasoning typically requires Chain-of-Thought prompting, which is accurate but incurs prohibitive latency and substantial test-time inferen

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

Robust-LLaVA: On the Effectiveness of Large-Scale Robust Image Encoders for Multi-modal Large Language Models

DGX agent

arXiv:2502.01576v2 Announce Type: replace Abstract: Multi-modal Large Language Models (MLLMs) excel in vision-language tasks but remain vulnerable to visual adversarial perturbations that can induce h

safetyarxiv-cs-cv
4 Jun 2026
Agents

ShareVerse: Multi-Agent Consistent Video Generation for Shared World Modeling

DGX agent

arXiv:2603.02697v2 Announce Type: replace-cross Abstract: This paper presents ShareVerse, a video generation framework enabling multi-agent shared world modeling, addressing the gap in existing works

agentsarxiv-cs-ai
4 Jun 2026
Agents

Stateful Visual Encoders for Vision-Language Models

DGX agent

arXiv:2606.04433v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in multi-image, multi-turn agentic settings where decisions depend on visual changes. However, in

agentsarxiv-cs-cl
4 Jun 2026
Model Releases

Video2LoRA: Parametric Video Internalization for Vision-Language Models

DGX agent

arXiv:2606.04351v1 Announce Type: cross Abstract: Processing video in vision-language models is expensive: each frame occupies hundreds of tokens, and inference cost scales with every frame and every

model-releasesarxiv-cs-cl
4 Jun 2026
Research

AI Model Extraction Attacks: Bypassing Single-Client Assumptions in Defenses

DGX agent

arXiv:2606.03381v1 Announce Type: cross Abstract: Ensuring the protection of Artificial Intelligence (AI) models deployed in military Command and Control (C2) systems and critical infrastructure is es

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Can Factual Opinions Be Edited (Manipulated) in Large Language Models?

DGX agent

arXiv:2606.03096v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly integrated into various domains, making knowledge editing techniques crucial yet potentially hazardous. Cu

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

CREward: A Type-Specific Creativity Reward Model

DGX agent

arXiv:2511.19995v2 Announce Type: replace Abstract: Creativity is a complex phenomenon. When it comes to representing and assessing creativity, treating it as a single undifferentiated quantity would

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Cryo-Bench: Benchmarking Foundation Models for Cryosphere Applications

DGX agent

arXiv:2603.01576v3 Announce Type: replace Abstract: Geo-Foundation Models (GFMs) have been evaluated across diverse Earth observation task including multiple domains and have demonstrated strong poten

model-releasesarxiv-cs-cv
3 Jun 2026
Research

dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching

DGX agent

arXiv:2506.06295v2 Announce Type: replace-cross Abstract: Autoregressive Models (ARMs) have long dominated the landscape of Large Language Models. Recently, a new paradigm has emerged in the form of d

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Generating the Modal Worker: A Cross-Model Audit of Race and Gender in LLM-Generated Personas Across 41 Occupations

DGX agent

arXiv:2510.21011v3 Announce Type: replace-cross Abstract: As generative AI tools are increasingly used to portray people in professional roles, understanding their racial and gender representational b

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

Patcher: Post-Hoc Patching of Backdoored Large Language Models

DGX agent

arXiv:2606.02995v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak backdoor attacks, where adversaries poison safety alignment data to embed hidden triggers that by

local-aiarxiv-cs-ai
3 Jun 2026
Model Releases

Pretraining Language Models on Historical Text

DGX agent

arXiv:2606.02991v1 Announce Type: cross Abstract: We introduce TypewriterLM, a 7.24B History language model (LM) trained exclusively on English text predating 1913. Developing History LMs requires add

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

TurtleAI: Benchmarking Multimodal Models for Visual Programming in Turtle Graphics

DGX agent

arXiv:2606.03626v1 Announce Type: cross Abstract: Vision-language models (VLMs) have been explored for visual programming, where they generate code to solve visual tasks. However, most prior work focu

model-releasesarxiv-cs-ai
3 Jun 2026
Applications

A Foundation Model for Wearable Movement Data in Mental Health Research

DGX agent

arXiv:2411.15240v5 Announce Type: replace-cross Abstract: Wearable movement data is collected by nearly all commercially available smartwatches and is a valuable resource for mental health research, r

applicationsarxiv-cs-ai
2 Jun 2026
Model Releases

Accuracy, Stability, and Repeated-Run Reliability of Large Language Models on Deterministic Programming Tasks

DGX agent

arXiv:2606.00920v1 Announce Type: cross Abstract: Run-level pass rate overstates retry-free coverage by up to 17.8 percentage points -- and the gap is largest precisely for mid-performing systems. We

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Active Exploring like a Pigeon: Reinforcing Spatial Reasoning via Agentic Vision-Language Models

DGX agent

arXiv:2606.02459v1 Announce Type: new Abstract: Enabling Vision-Language Models (VLMs) to perform spatial reasoning remains challenging. Existing approaches treat VLMs as passive observers, which is d

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Benchmarking Waitlist Mortality Prediction in Heart Transplantation Through Time-to-Event Modeling using New Longitudinal UNOS Dataset

DGX agent

arXiv:2507.07339v2 Announce Type: replace-cross Abstract: Decisions about managing patients on the heart transplant waitlist are currently made by committees of doctors who consider multiple factors,

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection

DGX agent

arXiv:2606.00101v1 Announce Type: cross Abstract: With the rapid advancement of artificial intelligence generated content (AIGC) technologies, video forgery has become increasingly prevalent, posing n

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents

DGX agent

arXiv:2606.02372v1 Announce Type: new Abstract: Equipping language agents with world models enables them to anticipate environment dynamics and evaluate candidate actions before execution. However, ex

safetyarxiv-cs-ai
2 Jun 2026
Research

EST-PRM: Stress-Testing Process Reward Models Before They Become Load-Bearing

DGX agent

arXiv:2606.00437v1 Announce Type: new Abstract: Process reward models (PRMs) are widely used in language-model training with dense step-level supervision. They assume PRM scores are stable proxies for

researcharxiv-cs-lg
2 Jun 2026
Safety

FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning

DGX agent

arXiv:2510.09222v3 Announce Type: replace Abstract: Flow Matching (FM) has shown remarkable ability in modeling complex distributions and achieves strong performance in offline imitation learning for

safetyarxiv-cs-lg
2 Jun 2026
Research

From Zero to Hero: Training-Free Custom Concept Spawning in World Models

DGX agent

arXiv:2606.02575v1 Announce Type: new Abstract: Autoregressive world models have emerged as a powerful paradigm for interactive video generation, allowing users to navigate dynamically generated envir

researcharxiv-cs-cv
2 Jun 2026
← Previous
1…8384858687…1030
Next →