AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Local Ai

Curia-MAE: Multi-Modal Multi-Anatomy MAE Pre-Training for 3D Medical Image Segmentation

DGX agent

arXiv:2608.05844v1 Announce Type: new Abstract: Radiology foundation models learn transferable representations that can be adapted to new tasks by training only small layers on top of a frozen encoder

local-aiarxiv-cs-cv
7 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Enhancing Anomaly Resilience in Research Networks: A Large-Scale Forecasting Benchmark for Dynamic Security Baselining

DGX agent

arXiv:2608.05605v1 Announce Type: cross Abstract: Research and Education Networks (RENs) serve as critical infrastructure for scientific discovery, yet they face a unique security paradox: their norma

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Enhancing Social Intelligence in LLMs with Hierarchical Reasoning and Utterance-Level Goal Rewarding

DGX agent

arXiv:2608.05832v1 Announce Type: new Abstract: Large language models (LLMs) excel in structured tasks but struggle with dynamic social interactions, where success requires long-term goal coordination

model-releasesarxiv-cs-cl
7 Aug 2026
Research

From Continuous Predictors to Clinical Thresholds: Early Evidence on Performance Trade-offs of Guideline-Based Categorisation for Ischaemic Stroke Outcome Prediction

DGX agent

arXiv:2608.05203v1 Announce Type: new Abstract: Machine learning models achieve strong predictive accuracy for 90-day outcome prediction in acute ischaemic stroke, yet clinical adoption is limited by

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Iterate or Widen? When Test-Time Refinement Helps LiDAR Scene Completion: A Controlled Study of Evidence Geometry, Training Coverage, and Compute

DGX agent

arXiv:2608.06014v1 Announce Type: new Abstract: Should a completion model spend extra test-time compute by iterating, or spend a similar parameter budget on a wider one-shot predictor? The answer is e

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

LLM Inference Under Bursty Workload Distribution: Modifying the WAIT Algorithm

DGX agent

arXiv:2608.06135v1 Announce Type: new Abstract: Large Language Models (LLMs) such as ChatGPT and Claude are widely used for information retrieval and problem-solving. Recent work has focused on improv

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

M^3R-Bench: A Unified Benchmark for Evidence-Grounded Multimodal Metaphor Understanding

DGX agent

arXiv:2608.05817v1 Announce Type: new Abstract: Metaphor enables the understanding of abstract concepts through cross-domain mappings while conveying affective attitudes. In multimodal scenarios, visu

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

MAC 2026: Advancing Micro-Action Analysis Towards Fine-Grained Understanding

DGX agent

arXiv:2607.16284v2 Announce Type: replace Abstract: Micro-Actions (MAs) are subtle and spontaneous human behaviors that provide important non-verbal cues in social interaction and affective communicat

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Operating Multi-Node Full Fine-Tuning on NVIDIA B300: A Field Report on Telemetry-Based Triage, Negative Results, and Operational Hardening

DGX agent

arXiv:2608.05944v1 Announce Type: cross Abstract: We report operational experience full-fine-tuning a 32.76B-parameter dense model (Qwen3-32B) on 16 x NVIDIA B300 (two nodes, FSDP / ZeRO-3) -- among t

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

SearchAuditor: Auditing and Attributing Failures in Long-Horizon Search Agents

DGX agent

arXiv:2608.05212v1 Announce Type: new Abstract: Deep search agents tackle challenging questions through long-horizon web interactions, a process that is both complex and fragile: small reasoning error

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Stochastic Parrots or Singing in Harmony? Testing Five Leading LLMs for their Ability to Replicate a Human Survey with Synthetic Data

DGX agent

arXiv:2603.00059v3 Announce Type: replace-cross Abstract: How well can AI-derived synthetic research data replicate the responses of human participants? An emerging literature has begun to engage with

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

TAU-Bench: From Anomaly Instance Tracking to Fine-Grained Video Anomaly Understanding

DGX agent

arXiv:2608.05699v1 Announce Type: new Abstract: Humans understand anomalous events through a coherent perceptual process in which they identify the focal instance, follow its behavior as the event unf

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

Visual Grounding in Zero-Shot Vision-Language Control

DGX agent

arXiv:2608.06154v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used as zero-shot controllers, but successful trajectories do not necessarily show that decisions are g

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

What Drives Test-Time Adaptation for CLIP? A Controlled Empirical Study from an Update Perspective

DGX agent

arXiv:2606.14299v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) such as CLIP have become a standard backbone for open-vocabulary recognition, yet their zero-shot predictions remain v

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Who Checks the Citations? Benchmarking Legal Hallucination Detection

DGX agent

arXiv:2606.21155v2 Announce Type: replace Abstract: Attorneys, judges, and pro se filers increasingly use AI to draft legal documents, yet these tools frequently fabricate citations. Despite predictio

model-releasesarxiv-cs-cl
7 Aug 2026
Research

A Mechanistic Analysis of Transformers for Dynamical Systems

DGX agent

arXiv:2512.21113v2 Announce Type: replace Abstract: Transformers are increasingly adopted for modeling and forecasting time-series, yet their internal mechanisms remain poorly understood from a dynami

researcharxiv-cs-lg
6 Aug 2026
Model Releases

A Survey of Agent Memory in the Second Half: Towards Self-Evolving and Long-Horizon Agents

DGX agent

arXiv:2602.06052v4 Announce Type: replace-cross Abstract: Research in artificial intelligence is shifting from model innovations and benchmark scores towards problem definition and rigorous real-world

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports

DGX agent

arXiv:2608.04682v1 Announce Type: cross Abstract: Coding agents powered by large language models (LLMs) are increasingly adopted in software engineering (SWE) scenarios, capable of fixing a specific b

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Adaptive Finite-Budget Training for CVaR Risk-Aware Q-Learning

DGX agent

arXiv:2608.04305v1 Announce Type: new Abstract: Risk-aware Q-learning (RaQL) provides a model-free, two-timescale estimator for dynamic risk objectives, but its finite-budget behavior remains fragile:

model-releasesarxiv-cs-lg
6 Aug 2026
Safety

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

DGX agent

arXiv:2608.05042v1 Announce Type: new Abstract: Leveraging pre-trained vision-language models (VLMs) to construct vision-language-action (VLA) models has emerged as a promising paradigm for 3D robot m

safetyarxiv-cs-ro
6 Aug 2026
Research

Deltoris: Enabling Real-time VLA Inference in Embodied AI via Bit-level Sparsity and Speculative Inference

DGX agent

arXiv:2608.04428v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have emerged as a key component in embodied AI. Among existing approaches, diffusion-based VLA models achieve supe

researcharxiv-cs-lg
6 Aug 2026
Model Releases

EgoAfford: Task-Oriented Affordance Grounding via Egocentric Referring Segmentation

DGX agent

arXiv:2608.04533v1 Announce Type: new Abstract: Part-level affordance grounding has advanced the localization of functional object regions associated with elemental actions. Extending this capability

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

LaPrune: Controllable Differentiable Sparsity at Million Scale

DGX agent

arXiv:2608.04057v1 Announce Type: cross Abstract: Top-k selection determines which components of a sparse model remain active. Hard selection blocks gradients, while continuous relaxations often coupl

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

MediRec: Enhancing Chinese Medication Recommendation with Explainable Clinical Reasoning

DGX agent

arXiv:2510.21084v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong potential for clinical decision support through their advanced language understanding and reaso

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

OmniRouting: A Semantic-Coupled Multimodal Benchmark for Constraint-Aware Spatial Reasoning in PCB Routing

DGX agent

arXiv:2608.04434v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in constraint-aware navigation, maze reasoning, and graph reasoning. However,

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

PhysMind: From Video to Executable Worlds for Training-Free Physical Reasoning

DGX agent

arXiv:2608.04575v1 Announce Type: cross Abstract: Reliable physical reasoning from video requires understanding how objects move, interact, and respond to interventions. Existing vision-language model

model-releasesarxiv-cs-ai
6 Aug 2026
Research

PSI3D: Plug-and-Play 3D Stochastic Inference with Slice-wise Latent Diffusion Prior

DGX agent

arXiv:2512.18367v2 Announce Type: replace-cross Abstract: Diffusion models are highly expressive image priors for Bayesian inverse problems. However, most diffusion models cannot operate on large-scal

researcharxiv-cs-lg
6 Aug 2026
Model Releases

RESPClinBench: Benchmarking Multimodal Clinical Decision-Making and Longitudinal Disease Management in Respiratory Specialty Care

DGX agent

arXiv:2608.04514v1 Announce Type: new Abstract: Background: Respiratory specialty care requires multimodal interpretation, longitudinal risk assessment, guideline-concordant intervention, and whole-co

model-releasesarxiv-cs-cl
6 Aug 2026
Research

REZE: Recognition-Based Zero-Shot Extraction for Video Temporal Grounding

DGX agent

arXiv:2608.04480v1 Announce Type: new Abstract: Video temporal grounding (VTG) refers to the task of identifying the time interval in a video that corresponds to a given natural-language query. A comm

researcharxiv-cs-cv
6 Aug 2026
Model Releases

Short-term load forecasting under EU-AI Act Requirements in Safety-Critical Environments: Results from a 41-day live challenge on the aggregated German transmission-grid load

DGX agent

arXiv:2608.05018v1 Announce Type: new Abstract: Short-term load forecasting (STLF) play a vital role in the electric power industry. It serves infrastructure that European and German law designate as

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Tactus: Open-Vocabulary Object Recognition from Low-Cost Pressure Arrays

DGX agent

arXiv:2608.04043v1 Announce Type: new Abstract: Resistive pressure arrays are the cheapest and most widely shipped tactile sensors, yet tactile representation learning has concentrated on optical sens

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

The Loss Does Not See the Basis, but Adam Does

DGX agent

arXiv:2608.05136v1 Announce Type: new Abstract: Gradient descent on a factored model W = UV^op is implicitly biased toward low-rank solutions, while Adam, starting from the same small initialization,

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Transfer Learning for Named Entity Recognition of Classical Latin through LLM Prompting

DGX agent

arXiv:2608.04015v1 Announce Type: new Abstract: With the increase in digitized resources of Classical Latin texts and modern breakthroughs of Large Language Models (LLMs), I contribute to ancient lang

model-releasesarxiv-cs-cl
6 Aug 2026
Safety

A Physics-Flavored Transformer Network for Parametrizing Contraction Dynamics of Engineered Skeletal Muscle Tissues

DGX agent

arXiv:2608.03927v1 Announce Type: new Abstract: Engineered Skeletal Muscle Tissues (ESMs) have become a key structure for biomedical disease modeling and pharmacological screening, yet their functiona

safetyarxiv-cs-lg
5 Aug 2026
Model Releases

Agents Catching Agents: Shortcut Cascades and Benchmark Gaming in Clinical Multi-Agent Systems

DGX agent

arXiv:2608.03744v1 Announce Type: new Abstract: Clinical decision support is moving toward committees of language-model agents deliberating on a shared workspace. We ask whether such committees can be

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Automated Visualization Code Synthesis via Multi-Path Reasoning and Feedback-Driven Optimization

DGX agent

arXiv:2502.11140v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become a cornerstone for automated visualization code generation, enabling users to create charts through na

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Beyond Simulations: What 20,000 Real Conversations Reveal About Mental Health AI Safety

DGX agent

arXiv:2601.17003v2 Announce Type: replace-cross Abstract: Mental-health AI safety is typically evaluated with small, simulation-based benchmarks that may not reflect the linguistic and contextual dive

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Beyond the Single Camera: Agentic Multi-View Reasoning in Sports Video Understanding

DGX agent

arXiv:2607.11844v2 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) achieve strong performance on single-view video understanding benchmarks. However, sports videos inv

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

CARE-Bench: Benchmarking Patient-Facing LLM Triage

DGX agent

arXiv:2608.03731v1 Announce Type: new Abstract: Patient-facing medical LLMs and agents increasingly answer symptom questions before clinician contact, where the key safety question is what action the

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

ConlangBench: Exploring Language Knowledge and Learning in LLMs through Diverse Constructed Languages

DGX agent

arXiv:2608.03505v1 Announce Type: new Abstract: Constructed languages (conlangs) are intentionally created human languages with a rich tradition of linguistic creativity. Despite their potential for s

model-releasesarxiv-cs-cl
5 Aug 2026
Research

EmbodiedVAE: Disentangled Video VAE for Efficient and Controllable Embodied Manipulation

DGX agent

arXiv:2608.02990v1 Announce Type: new Abstract: Latent diffusion models (LDMs) have recently significantly advanced embodied learning in constructing powerful embodied manipulation world models. Howev

researcharxiv-cs-ro
5 Aug 2026
Research

Every Wrong Answer Counts: Option-Level Psychometrics for LLM Multiple-Choice Benchmarks

DGX agent

arXiv:2608.02966v1 Announce Type: new Abstract: Most multiple-choice question (MCQ) benchmarks evaluate Large Language Models (LLMs) only by whether they select the correct answers. This binary scorin

researcharxiv-cs-cl
5 Aug 2026
Model Releases

HyVIC: A Metric-Driven Spatio-Spectral Hyperspectral Image Compression Architecture Based on Variational Autoencoders

DGX agent

arXiv:2603.26468v2 Announce Type: replace Abstract: The rapid growth of hyperspectral data archives in remote sensing (RS) necessitates effective compression methods for storage and transmission. Rece

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

ICO: Enhancing Semantic-Shift Jailbreaks via Iterative Context Optimization

DGX agent

arXiv:2608.03210v1 Announce Type: new Abstract: Foundation models have achieved remarkable success across diverse tasks, but they remain vulnerable. To investigate such vulnerabilities, semantic-shift

safetyarxiv-cs-cl
5 Aug 2026
Model Releases

Instruction Stacking Collapse: A Benchmark and the Capability-Dependent Value of Prompt Compilation

DGX agent

arXiv:2608.02639v1 Announce Type: cross Abstract: Production prompts rarely carry a single instruction. One system message may require valid JSON, a word limit, three citations, and a fixed tone at th

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Internalizing Academic Writing Workflows for Introduction Generation via Struct-Aware Policy Learning

DGX agent

arXiv:2608.03138v1 Announce Type: cross Abstract: Generating a rigorous paper introduction with large language models (LLMs) remains challenging, since it requires coordinating background, gap identif

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Inverted Detection and Control in Steering Vectors

DGX agent

arXiv:2608.02957v1 Announce Type: new Abstract: Steering vectors (SVs) are widely used to influence the expression of concepts (e.g., truthfulness) in large language model outputs. A key assumption un

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Knowing the Form, Not the Function: Automatically Auditing Answer--Authority Decoupling in Legal Benchmarks

DGX agent

arXiv:2608.02621v1 Announce Type: cross Abstract: Legal benchmarks typically score final answers even when models also state legal authority. We test whether answer correctness can serve as a proxy fo

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…393394395396397…1074
Next →