AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

Short-term load forecasting under EU-AI Act Requirements in Safety-Critical Environments: Results from a 41-day live challenge on the aggregated German transmission-grid load

DGX agent

arXiv:2608.05018v1 Announce Type: new Abstract: Short-term load forecasting (STLF) play a vital role in the electric power industry. It serves infrastructure that European and German law designate as

model-releasesarxiv-cs-ai
6 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SIGNPOST-Bench: Benchmarking Text-Vision Conflict Resolution in Multimodal Large Language Models

DGX agent

arXiv:2608.04244v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) make grounded predictions in real-world scenes by combining visual and textual cues, yet existing benchmarks

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

SimMOF: AI agent for Automated MOF Simulations

DGX agent

arXiv:2603.29152v2 Announce Type: replace Abstract: Metal-organic frameworks (MOFs) offer a vast design space, and as such, computational simulations play a critical role in predicting their structura

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses?

DGX agent

arXiv:2608.04828v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on skills, structured documents that specify when to act, which procedure to follow, and which tools

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

SmartMage: Dynamic Modality Orchestration for 3D Scene Understanding

DGX agent

arXiv:2608.05137v1 Announce Type: new Abstract: Understanding 3D scenes is fundamental to embodied intelligence, requiring joint reasoning over heterogeneous information from multiple modalities, incl

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

SpecDrop: Parameter-Free Category-Conditioned Routing for Modular Specialization

DGX agent

arXiv:2608.04084v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) networks pursue specialization through learned routers, gates, and load-balancing losses, yet at matched total-parameter budg

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Stabilizing Multi-Attack Adversarial Training via Bandit Optimization

DGX agent

arXiv:2511.12265v2 Announce Type: replace-cross Abstract: Deep Neural Networks (DNNs) remain vulnerable to diverse adversarial perturbations, motivating multi-attack adversarial training (AT) for impr

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Strategic Evaluation of Planning Strategies for LLM Agents in Cyber-Physical Systems

DGX agent

arXiv:2608.04265v1 Announce Type: cross Abstract: Evaluations of LLM planning agents largely ask whether a task succeeds or a declared plan is followed. In strategic cyber-physical systems, a stronger

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Strengthening Target-Language Features: SAE-Based Steering for Multilingual Inference

DGX agent

arXiv:2608.04904v1 Announce Type: new Abstract: Multilingual large language models exhibit substantial performance differences across languages, while existing adaptation methods often require paramet

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

STRIVE: Probing Reasoning Limits in Graded Plausibility Generation and Evaluation

DGX agent

arXiv:2608.04567v1 Announce Type: new Abstract: Event knowledge concerns who does what to whom. Psycholinguists use event-plausibility judgments to examine how this knowledge supports human language p

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Tactus: Open-Vocabulary Object Recognition from Low-Cost Pressure Arrays

DGX agent

arXiv:2608.04043v1 Announce Type: new Abstract: Resistive pressure arrays are the cheapest and most widely shipped tactile sensors, yet tactile representation learning has concentrated on optical sens

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching

DGX agent

arXiv:2608.04568v1 Announce Type: new Abstract: As a key capability for embodied intelligence, 3D visual grounding (3DVG) has been predominantly studied in indoor scenes with RGB-D or point-cloud inpu

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Teaching Foundation Models to Read mmWave: Pose-Guided Kinematic Representation for Human Behavior Understanding

DGX agent

arXiv:2608.04127v1 Announce Type: new Abstract: Large language model agents need to perceive human behavior in physical environments. Millimeter-wave (mmWave) radar provides a privacy-friendly and con

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains

DGX agent

arXiv:2608.05138v1 Announce Type: cross Abstract: Modern Greek is absent from NVIDIA's Nemotron retrieval models and from major multilingual retrieval benchmarks, despite being important for retrieval

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Test, then Route: How Language Models Execute In-Context Conditional Rules Across Models and Languages

DGX agent

arXiv:2608.04183v1 Announce Type: new Abstract: When a language model follows an in-context conditional rule such as 'if P(x) then A else B,' does it assemble a runtime circuit with one module that te

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Text2GraphQuery-Bench: A Text to Graph Query Benchmark

DGX agent

arXiv:2602.11745v2 Announce Type: replace Abstract: Graph models are fundamental to data analysis in domains rich with complex relationships. Unlike SQL, which benefits from a rel- atively unified sta

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

The Calibration Floor: Format Repair Can Masquerade as Self-Correction at Small-to-Mid Scale

DGX agent

arXiv:2608.04355v1 Announce Type: new Abstract: Accuracy changes after language-model self-revision are usually interpreted as changes in reasoning. We show this can fail at the answer-extraction boun

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

The Evaluator Is Part of the Experiment: Measuring Open-Ended LLM Conformity

DGX agent

arXiv:2608.04463v1 Announce Type: new Abstract: Prior work on LLM conformity largely measures discrete answer flips under verifiable labels. Open-ended revisions require a different measurement strate

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

The First EgoCross Challenge at EgoVis 2026: Cross-Domain Egocentric Video Question Answering

DGX agent

arXiv:2608.04589v1 Announce Type: cross Abstract: EgoCross is a cross-domain egocentric video question answering benchmark designed to evaluate whether multimodal large language models can generalize

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

The Loss Does Not See the Basis, but Adam Does

DGX agent

arXiv:2608.05136v1 Announce Type: new Abstract: Gradient descent on a factored model W = UV^op is implicitly biased toward low-rank solutions, while Adam, starting from the same small initialization,

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

The Order Is the Guarantee: Verifier-Budgeted Code Deletion with Static-First Learned Proposals

DGX agent

arXiv:2608.04611v1 Announce Type: cross Abstract: Frontier coding models now match or exceed strong human reference points on programming benchmarks, yet benchmark success does not imply maintainable

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

The Yokai Learning Environment: Tracking Beliefs Over Space and Time

DGX agent

arXiv:2508.12480v3 Announce Type: replace Abstract: The ability to cooperate with unknown partners is a central challenge in cooperative AI and widely studied in the form of zero-shot coordination (ZS

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Thinking with Anchors: Grounded and Efficient Document Reasoning

DGX agent

arXiv:2608.04424v1 Announce Type: new Abstract: Existing document understanding benchmarks have largely focused on locating page elements, yet real-world document intelligence requires models to reaso

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

TIDE: A Physically Diverse 3D Turbulence Benchmark Dataset for Advancing Scientific Machine Learning

DGX agent

arXiv:2608.04222v1 Announce Type: cross Abstract: Turbulence is a central testbed for machine learning on physical dynamics because its governing laws are known exactly. However, most existing studies

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Toward Federated Large Language Models in Medicine: A Parameter-Efficient Framework for Privacy-Preserving, Multi-Institutional Adaptation

DGX agent

arXiv:2601.22124v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly adapted for medical applications, but most are trained using data from a single institution because pr

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Toward Integrating Adaptive Experience Replay and Online Uncertainty Estimation in Safe Actor-Critic Optimal Control

DGX agent

arXiv:2608.04732v1 Announce Type: cross Abstract: Safe actor-critic control often treats barrier filtering, uncertainty estimation, and experience replay as separate modules, even though each changes

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning

DGX agent

arXiv:2608.05139v1 Announce Type: new Abstract: Long-horizon reasoning in recent LLMs demands that the model switch between distinct skills inside a reasoning chain, such as first doing a math derivat

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Towards a satellite image manipulation and deepfake localization benchmark dataset

DGX agent

arXiv:2608.04840v1 Announce Type: cross Abstract: Verifying the authenticity of satellite imagery has become increasingly critical given advances in generative artificial intelligence. Highly realisti

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Towards Trustworthy Hypergraph Neural Networks under Label Noise

DGX agent

arXiv:2608.04377v1 Announce Type: cross Abstract: Hypergraph neural networks (HGNNs) have demonstrated remarkable capabilities in processing complex higher-order relationships. However, their performa

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Training Crossroads for Recurrent Vision Transformers: Recurrence, Neural ODEs, and Deep Supervision

DGX agent

arXiv:2608.04879v1 Announce Type: cross Abstract: Vision Transformers (ViTs) achieve strong image-recognition performance, but their parameter count grows linearly with depth when each block is indepe

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Transfer Learning for Named Entity Recognition of Classical Latin through LLM Prompting

DGX agent

arXiv:2608.04015v1 Announce Type: new Abstract: With the increase in digitized resources of Classical Latin texts and modern breakthroughs of Large Language Models (LLMs), I contribute to ancient lang

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Trident : How to Break Deep Reinforcement Learning Cyber Defenses (Agentic)

DGX agent

arXiv:2608.04317v1 Announce Type: cross Abstract: Autonomous cyber defense systems based on Deep Reinforcement Learning (DRL) have attracted significant research attention, yet remain evaluated almost

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Tropical Algebraic Geometry for Neuronal Representations: An Arakelov-Green Measure Based Descriptor for Graph Learning

DGX agent

arXiv:2608.04460v1 Announce Type: cross Abstract: The quantitative analysis of 3D neuronal morphologies requires capturing both graph topology and spatial geometry. Current message-passing Graph Neura

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Two-Level Meta-Rubrics for Evaluating Open-Ended Generation: GAMUT, a Benchmark for Factual Completeness

DGX agent

arXiv:2607.19322v2 Announce Type: replace Abstract: Rubric-based evaluation of open-ended generation faces a fundamental tension between expressiveness and reliability. Authoring a faithful rubric req

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

UG-UMRE: Uncertainty-Guided Modality Augmentation and Distributional Calibration for Unified Multimodal Relation Extraction

DGX agent

arXiv:2608.04949v1 Announce Type: cross Abstract: Unified Multimodal Relation Extraction (UMRE) aims to identify intra-modal and cross-modal relations between textual entities and visual objects. Howe

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models

DGX agent

arXiv:2608.04701v1 Announce Type: new Abstract: The abundance of casually captured monocular videos and images on social media provides a valuable source for immersive content creation, where generati

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation

DGX agent

arXiv:2608.04902v1 Announce Type: new Abstract: Video-to-audio (V2A) generation extends image-to-audio generation (I2A) by introducing consecutive frames that provide essential temporal cues for audio

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

What We Observe as LLM Behavior Can Be a Side-effect of Inference Backend

DGX agent

arXiv:2608.04714v1 Announce Type: cross Abstract: Benchmark scores are reported as properties of a model, yet the inference framework used to produce them, such as HuggingFace, vLLM, or Ollama, are co

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs

DGX agent

arXiv:2608.04893v1 Announce Type: cross Abstract: Multi-agent LLM systems relay key--value caches instead of text and credit their gains to exchanged ``latent thoughts''. That credit is a claim about

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning

DGX agent

arXiv:2608.04726v1 Announce Type: new Abstract: Multimodal large language models increasingly reason over screenshots and documents where the task itself may be written in pixels. Yet benchmarks usual

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Why Ranking Anomaly Detection Algorithms Isn't as Reliable as You May Think

DGX agent

arXiv:2608.04613v1 Announce Type: new Abstract: Anomaly detection is a safety-critical machine learning problem with applications ranging from fraud detection to network intrusion prevention and indus

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models

DGX agent

arXiv:2608.04964v1 Announce Type: new Abstract: Interactive video world models are essential for long-horizon planning and exploration, yet they suffer from compounding errors. Post-training methods s

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Zero-shot reasoning for simulating scholarly peer-review

DGX agent

arXiv:2510.02027v2 Announce Type: replace Abstract: Scholarly publishing requires scalable scrutiny supported by auditable evidence. This paper presents a two-component benchmark of xPeer, the peer-re

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

3DGSI-Assessor: A Large-Scale Dataset and An LMM-based Method for 3D Gaussian Splatting Image Quality Assessment

DGX agent

arXiv:2608.03279v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has become a dominant representation for real-time novel view synthesis (NVS), yet its storage footprint makes compression

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

A Unified 2D Framework for DeepLesion Detection, Segmentation and Short Report Generation

DGX agent

arXiv:2608.02805v1 Announce Type: cross Abstract: In previous work, we integrated large language models (LLMs) into the lesion segmentation model based on the ULS23 DeepLesion dataset, using short-for

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models

DGX agent

arXiv:2608.03112v1 Announce Type: cross Abstract: Vision-language models excel at image and video understanding but suffer from high inference latency due to the need to process thousands of tokens pe

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Adversarial Fast-Moving Real-World Domains as Test Beds for Benchmarking AI Scientist Capabilities

DGX agent

arXiv:2608.03569v1 Announce Type: new Abstract: Benchmarking the ability of AI scientists to generate novel ideas is notoriously difficult. Existing benchmarks in this field have made progress in eval

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation

DGX agent

arXiv:2608.03166v1 Announce Type: new Abstract: Role-Playing Language Agents (RPLAs) are increasingly deployed in high-stakes applications such as healthcare assistance, customer support, and educatio

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…2425262728…357
Next →