AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,001 results
Model Releases

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback

DGX agent

arXiv:2607.26094v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard approach for aligning large language models with human preferences, but its quality

model-releasesarxiv-cs-cl
30 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

MetaKoopman: Bayesian Meta-Learning of Koopman Operators for Modeling Structured Dynamics under Distribution Shifts

DGX agent

arXiv:2607.26345v1 Announce Type: new Abstract: Modeling and forecasting nonlinear dynamics under distribution shifts is essential for robust decision-making in real-world systems. In this work, we pr

agentsarxiv-cs-lg
30 Jul 2026
Agents

MindForge: Teaching Small Language Models Whole-Life-Cycle Software Engineering via Source-Free Program Synthesis

DGX agent

arXiv:2607.27146v1 Announce Type: cross Abstract: Coding agents have made substantial progress on software engineering tasks that modify existing codebases, including bug fixing and feature implementa

agentsarxiv-cs-cl
30 Jul 2026
Research

ReDiSC: A Reparameterized Masked Diffusion Model for Scalable Node Classification with Structured Predictions

DGX agent

arXiv:2507.14484v2 Announce Type: replace Abstract: In recent years, graph neural networks (GNN) have achieved unprecedented successes in node classification tasks. Although GNNs inherently encode spe

researcharxiv-cs-lg
30 Jul 2026
Applications

RVC-NMPC: Nonlinear Model Predictive Control with Reciprocal Velocity Constraints for Mutual Collision Avoidance in Agile UAV Flight

DGX agent

arXiv:2512.08574v2 Announce Type: replace Abstract: This paper presents an approach to mutual collision avoidance based on Nonlinear Model Predictive Control (NMPC) with time-dependent Reciprocal Velo

applicationsarxiv-cs-ro
30 Jul 2026
Model Releases

ToxScreen: Detecting Whether an LLM Has Been Poisoned

DGX agent

arXiv:2607.26849v1 Announce Type: cross Abstract: As large language models (LLMs) are deployed in high-stakes domains, adversaries may poison training data to implant backdoors: hidden triggers that c

model-releasesarxiv-cs-lg
30 Jul 2026
Hardware

Visko Orbis 1.0: A Live Model for Real-Time Interactive Long Video Generation

DGX agent

arXiv:2607.26694v1 Announce Type: new Abstract: We present Visko Orbis 1.0, a Live Model for real-time, interactive long-video generation. Users can change the prompt at any moment during generation,

hardwarearxiv-cs-cv
30 Jul 2026
Model Releases

What is the best intelligence/stable model currently for a single GB10/DGX spark?

DGX agent

Is Qwen 3.6 27b still the go' ol' reliable at this point? I know 35b is faster but it just doesn't give as good results. Is it possible to run deepseek v4 flash on a single spark at decent tk/s withou

model-releasesr-localllama
30 Jul 2026
Research

A Study of Crosslinguistic Influence in Language Models

DGX agent

arXiv:2601.21587v2 Announce Type: replace Abstract: The sequential acquisition of languages inevitably leads to Crosslinguistic Influence (CLI), where the syntactic properties of a first language (L1)

researcharxiv-cs-cl
29 Jul 2026
Model Releases

Conformal Cascade: Distribution-Free Accuracy Guarantees for Multi-Tier LLM Inference

DGX agent

arXiv:2607.25018v1 Announce Type: new Abstract: Large language model (LLM) cascades reduce inference cost by routing easy queries to a small model and deferring hard queries to a larger one. Productio

model-releasesarxiv-cs-lg
29 Jul 2026
Safety

Construction-Driven Injection: Linguistically-Grounded Edit-Based Code-Mixing Fingerprints for Large Language Models

DGX agent

arXiv:2607.25633v1 Announce Type: cross Abstract: Large language models (LLMs) are costly intellectual assets that remain exposed to unauthorized redistribution and commercial misuse. Injected fingerp

safetyarxiv-cs-ai
29 Jul 2026
Local Ai

DC-WAM: Dynamic-Centric Visual Supervision and Reasoning for World-Action Models

DGX agent

arXiv:2607.25918v1 Announce Type: new Abstract: World-Action Models (WAMs) augment robot policies with future visual prediction, but it remains unclear what the visual modality should learn for contro

local-aiarxiv-cs-ro
29 Jul 2026
Research

Diff-ID: Identity Consistent Facial Image Generation and Morphing via Diffusion Models

DGX agent

arXiv:2607.25078v1 Announce Type: new Abstract: Generative diffusion models have revolutionized facial image synthesis, yet robust identity preservation in high resolution outputs remains a critical c

researcharxiv-cs-cv
29 Jul 2026
Research

Evaluating Communicative Belief Updates in Large Language Models via Implicature Recognition and Cancellation

DGX agent

arXiv:2607.25094v1 Announce Type: cross Abstract: Human language is driven by unspoken beliefs and belief updates, making these critical to model for successful communication between large language mo

researcharxiv-cs-ai
29 Jul 2026
Local Ai

INTACT: Isomorphic Intent-to-Action Learning for Search-Free World Models

DGX agent

arXiv:2607.26056v1 Announce Type: new Abstract: Forward latent world models predict how actions change a scene, but recover actions for a desired change only through expensive test-time search. We int

local-aiarxiv-cs-ro
29 Jul 2026
Research

Lantern: Conflict-Aware Gradient Blending for Physics-Guided Diffusion Models in Calorimeter Simulation

DGX agent

arXiv:2607.25060v1 Announce Type: cross Abstract: Monte Carlo simulation of calorimeter showers is a principal bottleneck for the High-Luminosity LHC, and diffusion models have emerged as fast, high-f

researcharxiv-cs-ai
29 Jul 2026
Safety

LLM as Forecasting Planner: Training-Free Text Conditioning for Time-Series Foundation Models

DGX agent

arXiv:2607.24892v1 Announce Type: cross Abstract: Text-conditioned time-series forecasting predicts a series from both its numerical history and natural-language context, allowing forecasts to account

safetyarxiv-cs-ai
29 Jul 2026
Safety

Real-Time Driver Safety Scoring Through Inverse Crash Probability Modeling

DGX agent

arXiv:2603.14841v3 Announce Type: replace-cross Abstract: Road crashes remain a leading cause of preventable fatalities. Existing prediction models predominantly produce binary outcomes, which offer l

safetyarxiv-cs-ai
29 Jul 2026
Research

Unified Semantic Modeling Framework for Large-Scale Job Understanding at LinkedIn

DGX agent

arXiv:2607.24783v1 Announce Type: new Abstract: Job understanding is critical to LinkedIn's mission of connecting talent with opportunity. This task involves transforming unstructured and noisy job po

researcharxiv-cs-ai
29 Jul 2026
Applications

We hope these experiments serve as a reminder that evals rarely measure models in isolation—they also measure a bundle of less visible choic…

DGX agent

We hope these experiments serve as a reminder that evals rarely measure models in isolation—they also measure a bundle of less visible choices about API settings, harness design, and prompting. If you

applicationsopenai--x
29 Jul 2026
Research

Benchmarking the Domain Gap: Model Selection Instability Under Domain Shift in Video Capsule Endoscopy

DGX agent

arXiv:2607.22736v1 Announce Type: new Abstract: Video capsule endoscopy (VCE) classification is typically evaluated within a single dataset, yet clinical deployment demands robustness across acquisiti

researcharxiv-cs-cv
28 Jul 2026
Research

CAPT: A Multi-task Continuous Autoregressive Transformer enabling Cross-dataset and Cross-species Transfer for Calcium Population Dynamics

DGX agent

arXiv:2607.23258v1 Announce Type: new Abstract: Large-scale calcium imaging has created an opportunity to build foundation-style models for neural population dynamics, but a central question remains u

researcharxiv-cs-ai
28 Jul 2026
Research

DSTFView: Multi-View Cloud-Edge Workload Forecasting with Dual-Input Spatio-Temporal-Frequency Modeling

DGX agent

arXiv:2607.22565v1 Announce Type: new Abstract: With the widespread deployment of edge-side AI inference, edge platforms are increasingly required to support latency-sensitive, highly concurrent, and

researcharxiv-cs-ai
28 Jul 2026
Model Releases

ESF-Bench: Benchmarking Challenging Slot-Filling Scenarios for Real-World Enterprise Applications

DGX agent

arXiv:2607.23326v1 Announce Type: new Abstract: The rapid rise of large language models (LLMs) has driven transformative adoption across enterprises. However, deploying these models in real-world sett

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

FlowCTS: On-policy Continuous Trajectory Supervision of Flow Models

DGX agent

arXiv:2607.24522v1 Announce Type: cross Abstract: While on-policy distillation (OPD) effectively addresses sparse rewards and exposure bias in large language model post-training, its extension to flow

safetyarxiv-cs-cv
28 Jul 2026
Local Ai

Let AI Agents Translate Networks, Not Reason About Them

DGX agent

arXiv:2607.22947v1 Announce Type: new Abstract: A formal model enables verifying reachability, localizing an outage, or anticipating the blast radius of a change. Yet, virtually no production network

local-aiarxiv-cs-ai
28 Jul 2026
Agents

Lexical discovery in unknown environments orchestrated by Large Language Models

DGX agent

arXiv:2607.22591v1 Announce Type: new Abstract: Populations of autonomous agents deployed in unknown environments (e.g. planetary or deep-sea exploration) must develop shared vocabularies to refer to

agentsarxiv-cs-ai
28 Jul 2026
Research

Logit-Coordinate Generative Models for Mixed Continuous-Categorical Tabular Data

DGX agent

arXiv:2607.23348v1 Announce Type: cross Abstract: Mixed continuous--categorical data pose a representation problem for continuous generative models. Flow Matching and Gaussian diffusion operate in Euc

researcharxiv-cs-lg
28 Jul 2026
Model Releases

MAViE: A Multi-scale Adaptive Vision Encoder for Fine-grained Visual Perception and Efficient Multimodal Reasoning

DGX agent

arXiv:2607.24424v1 Announce Type: new Abstract: Vision-language models commonly project all tokens produced by a pretrained vision encoder into a large language model. However, final-layer features ca

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

MoLGE: Mixture of Language Group Experts for Efficient Scaling of Massively Multilingual Speech Recognition

DGX agent

arXiv:2607.24030v1 Announce Type: new Abstract: Massively multilingual automatic speech recognition (ASR) models covering hundreds of languages must maintain robust performance across diverse linguist

model-releasesarxiv-cs-cl
28 Jul 2026
Applications

Music-Source-Separation-Training (MSST): A Unified Framework for Training and Evaluating Music Demixing Models

DGX agent

arXiv:2607.23395v1 Announce Type: cross Abstract: Music Source Separation (MSS), the task of recovering individual sound components (stems) from a polyphonic mixture, is central to applications rangin

applicationsarxiv-cs-lg
28 Jul 2026
Safety

N_0-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens

DGX agent

arXiv:2607.23782v1 Announce Type: new Abstract: We present N_0-VTLA, a vision-tactile-language-action (VTLA) foundation model capable of (1) fine-grained contact-rich manipulation with tactile percept

safetyarxiv-cs-ro
28 Jul 2026
Research

Reason-Mediated Behavioral Models for Auditing LLM Social Simulators

DGX agent

arXiv:2607.24649v1 Announce Type: new Abstract: Large language models are increasingly used as social simulators, including as synthetic survey respondents. Most evaluations ask whether simulated outc

researcharxiv-cs-ai
28 Jul 2026
Safety

Structured Redundancy Modeling for Efficient Visual Token Pruning in High-Resolution MLLMs

DGX agent

arXiv:2607.23046v1 Announce Type: new Abstract: Recent high-resolution Multimodal Large Language Models (MLLMs) generate thousands of visual tokens per input, leading to a visual token explosion that

safetyarxiv-cs-cv
28 Jul 2026
Model Releases

TLA^{+}-Bench: An Execution-Grounded Benchmark and Dataset for Natural-Language to TLA+ Specification Generation

DGX agent

arXiv:2607.23425v1 Announce Type: cross Abstract: Large language models increasingly write TLA^{+} formal specifications from natural-language descriptions, but progress is hard to measure: existing r

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Toward Automated Detection of Documentation Inconsistencies in Electronic Health Records

DGX agent

arXiv:2607.22954v1 Announce Type: new Abstract: Objective: To characterize the kinds of internal documentation inconsistencies a general-domain large language model (LLM) can surface from real-world d

model-releasesarxiv-cs-cl
28 Jul 2026
Safety

Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex

DGX agent

arXiv:2607.23333v1 Announce Type: cross Abstract: We revisit the regret loss framework introduced in Park et al. (2025), which uses decision-theoretic regret as a direct loss function for training mod

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Understanding Machine Unlearning Through the Lens of Mode Connectivity

DGX agent

arXiv:2607.23970v1 Announce Type: cross Abstract: Machine Unlearning aims to remove undesired information from trained models without full retraining from scratch. Despite recent progress, the loss la

model-releasesarxiv-cs-ai
28 Jul 2026
Applications

Visual Information Extraction from Documents via Classification-Guided Large Vision-Language Models

DGX agent

arXiv:2607.22723v1 Announce Type: new Abstract: Visual information extraction (VIE) from visually rich documents remains challenging due to high layout variability and real-world impairments. Existing

applicationsarxiv-cs-cv
28 Jul 2026
Model Releases

Zing: Social Mind for LLMs

DGX agent

arXiv:2607.23740v1 Announce Type: new Abstract: As large language models move from isolated task solving toward long-term service in human environments, they require social intelligence: the ability t

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Be Consistent! Enhancing Robust Visual Reasoning in LVLMs with Consistency Constraints

DGX agent

arXiv:2607.21722v1 Announce Type: new Abstract: While Large Vision-Language Models (LVLMs) exhibit strong perceptual capabilities, they remain vulnerable in visual reasoning tasks. Existing benchmarks

model-releasesarxiv-cs-cv
27 Jul 2026
Research

Diffusion Models in Medical Image Inpainting: Challenges, Solution Taxonomy, and Future Directions

DGX agent

arXiv:2607.21904v1 Announce Type: cross Abstract: Image inpainting aims to reconstruct missing or corrupted regions of an image while preserving as much as possible, visual and semantic consistency. I

researcharxiv-cs-cl
27 Jul 2026
Model Releases

Pixels for Programs? A Cross-Provider Case Study of Input-Token Accounting for Source Code as Text and Images

DGX agent

arXiv:2607.21672v1 Announce Type: cross Abstract: Long source-code contexts consume many text tokens, motivating the proposal to render code as images for vision-language models. Recent work asks whet

model-releasesarxiv-cs-cv
27 Jul 2026
Hardware

Six Agent Harness Capabilities for Higher Model Performance

DGX agent

The performance of AI agents depends not only on the underlying models but also heavily on their “harness”—the surrounding architecture that supplies context, state management, action execution, and t

hardwarenvidia-developer
27 Jul 2026
Model Releases

Open-weight 4B models approach o3-level medical question answering in Swedish [P]

DGX agent

I have been running some experiments with smaller open-weight LLMs on multiple-choice questions of Swedish medical licensing exams. On a dataset called MedQA-SWE, GPT-4 scored 84% accuracy in 2024 and

model-releasesr-machinelearning
26 Jul 2026
Applications

Attention-based Experience Replay Framework for Continual Learning of Agnostic Time Series Forecasting Models

DGX agent

arXiv:2607.20493v1 Announce Type: new Abstract: Deep learning has led to remarkable progress in artificial intelligence, particularly in robotics, imaging and sound processing. However, a major limita

applicationsarxiv-cs-ai
24 Jul 2026
Model Releases

[audio.cpp] Release 0.4: Higgs Audio v3 TTS 4B (10x real time)+ Fish Audio S2 Pro in C++/GGML, full GGUF loading, Q8 speed and VRAM gains

DGX agent

audio.cpp again :) Release 0.4 is out. The headline this time is new high-quality TTS coverage plus GGUF becoming a first-class across the project. What’s new: Added Higgs Audio v3 TTS 4B, Fish Audio

model-releasesr-localllama
24 Jul 2026
Model Releases

Break Through the Compression Bottleneck: From Theory to Practice

DGX agent

arXiv:2607.20434v1 Announce Type: cross Abstract: As the parameter size of language models continues to grow, effective model compression is required to reduce their computational and memory overhead.

model-releasesarxiv-cs-ai
24 Jul 2026
← Previous
1…216217218219220…1271
Next →