AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

GPUSimBench: Towards Scalable and Reliable GPU-Accelerated Simulators in Embodied AI

DGX agent

arXiv:2607.13059v1 Announce Type: new Abstract: Data-driven embodied AI is rapidly transitioning into a paradigm that scales training through massively parallel simulation, where GPU-accelerated simul

model-releasesarxiv-cs-ro
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Graded Entity-Familiarity Readouts in Language Models: Polish Adaptation, Cross-Language Robustness, and Refusal Steering

DGX agent

arXiv:2607.13568v1 Announce Type: cross Abstract: Can a language model estimate its familiarity with an entity before generating an answer? We study activations at the final prompt token in twelve ins

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

Heavy-Tailed Flow Matching via Random Clocks

DGX agent

arXiv:2607.13841v1 Announce Type: new Abstract: Heavy-tailed data arise in many domains where rare events carry disproportionate importance, such as imbalanced image datasets, financial returns, and w

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

HEDGEHOG: Hierarchical Evaluation of Drug Generators Through Rigorous Filtration

DGX agent

arXiv:2607.13155v1 Announce Type: new Abstract: Generative molecular models can support early drug discovery by proposing new candidate compounds de novo. In practice, useful candidates must balance t

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

How Far Can Root Cause Analysis Go on Real-World Telemetry Data?

DGX agent

arXiv:2607.13548v1 Announce Type: new Abstract: Identifying root causes in production microservice failures requires reasoning over large-scale, multimodal telemetry spanning metrics, logs, and traces

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

HRIBench: Benchmarking Interaction-Centric Human-Robot Collaboration

DGX agent

arXiv:2607.13056v1 Announce Type: cross Abstract: Current vision-language-action (VLA) benchmarks primarily evaluate isolated manipulation skills while leaving human-robot interaction structure largel

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

Implementations of Quantum and Classical Topology-Aligned Architectures for Molecular Property Prediction

DGX agent

arXiv:2607.13737v1 Announce Type: new Abstract: For low-data and resource-constrained regimes typical of quantum chemistry, parameter-efficient learning is a key objective. Here, we propose a topology

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models

DGX agent

arXiv:2607.13408v1 Announce Type: cross Abstract: Recent text-to-audio models generate high-quality audio, but often fail to follow instructions involving multiple sound events and temporal order. Thi

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Industrial Dexterity Benchmark: A Hardware-Software Benchmarking Platform for Industrial Dexterous Manipulation

DGX agent

arXiv:2607.14021v1 Announce Type: new Abstract: Dexterous manipulation remains a critical bottleneck in industrial automation; tasks such as cable routing, connector insertion, and precision assembly

model-releasesarxiv-cs-ro
16 Jul 2026
Model Releases

Inference Economics of Enterprise Coding Agents: A Case Study of Cloud vs. On-Premise LLMs

DGX agent

arXiv:2607.13080v1 Announce Type: cross Abstract: Autonomous coding agents force engineering organizations to choose between API-based frontier models -- strong reasoning at high token cost -- and on-

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate

DGX agent

arXiv:2510.10002v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in sensitive everyday contexts -- offering personal advice, mental health support, and mor

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Interventional Grounding Audits: Black-Box Premise-Dependency Tests for LLM Chain-of-Thought via Predicate Substitution

DGX agent

arXiv:2607.13069v1 Announce Type: new Abstract: Large language models produce chain-of-thought (CoT) reasoning that appears logically sound yet may not genuinely depend on its stated premises. We intr

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Is the Statistical Advantage Worth the Cost? An Empirical Comparison of KANs and MLPs for Structured Data Classification

DGX agent

arXiv:2607.13413v1 Announce Type: cross Abstract: This study presents an empirical benchmarking comparison between Kolmogorov-Arnold Networks (KANs) and Multi-Layer Perceptrons (MLPs) on structured ta

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Learning Robust Execution in Robotic Manipulation with Agentic Reinforcement Learning

DGX agent

arXiv:2607.13818v1 Announce Type: new Abstract: Robotic manipulation poses fundamental challenges due to uncertainty, long-horizon execution, and compounding errors, which can easily destabilize execu

model-releasesarxiv-cs-ro
16 Jul 2026
Model Releases

LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents

DGX agent

arXiv:2607.13041v1 Announce Type: cross Abstract: Large Language Model (LLM) based AI educational content generation systems are increasingly being developed, yet no standardised benchmark exists to s

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Lighthouse RL: Sample-Efficient Circuit Optimization via Strategic Reset Points

DGX agent

arXiv:2607.14008v1 Announce Type: new Abstract: In this paper, we introduce Lighthouse RL, a sample-efficient reinforcement learning (RL) approach for analog circuit sizing. Traditional methods lack g

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

Look Again Before You Abstain:Budgeted Conformal Evidence Acquisition for Reliable Vision-Language Model

DGX agent

arXiv:2606.16667v2 Announce Type: replace Abstract: Large vision-language models (LVLMs) hallucinate: they assert visual details that the image does not support. A principled remedy is selective predi

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

M+Adam: Low-Precision Training via Additive-Multiplicative Optimization

DGX agent

arXiv:2607.10611v2 Announce Type: replace Abstract: Training with quantized weights can reduce costs but often results in degraded accuracy, especially when optimization is carried out in low precisio

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation

DGX agent

arXiv:2607.09142v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in online medical consultation, yet existing benchmarks remain poorly aligned with real clini

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Merging Reaction to Cognition: A Hybrid Cognitive Strategy for Odour Source Localisation in Natural Environments

DGX agent

arXiv:2607.13853v1 Announce Type: new Abstract: Chemical pollutants released into the environment are transported by turbulent flows, generating complex, intermittent plume structures that threaten ec

model-releasesarxiv-cs-ro
16 Jul 2026
Model Releases

Mixed-Timescale Differential Coding for Downlink Model Broadcast in Wireless Federated Learning

DGX agent

arXiv:2607.13119v1 Announce Type: cross Abstract: In standard federated learning systems, the parameter server broadcasts the global model to the participating devices in every iteration. Motivated by

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

Mono-Z Dark Matter Search with Neural Spline Flows Using CMS Run 2015D Open Data

DGX agent

arXiv:2607.13771v1 Announce Type: new Abstract: We report a search for dark matter (DM) produced in association with a leptonically decaying (Z) boson at (sqrt{s}=13) TeV using CMS Run 2015D open data

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

MxGPS: Multiplex Graph Transformers for a Power Grid Foundation Model

DGX agent

arXiv:2607.13763v1 Announce Type: cross Abstract: Single-task fine-tuning of graph neural networks (GNNs) for power grid problems exhibits a systematic failure mode: models that achieve the lowest in-

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

NeMo: Needle in a Montage for Video-Language Understanding

DGX agent

arXiv:2509.24563v3 Announce Type: replace Abstract: Recent advances in video large language models (VideoLLMs) call for new evaluation protocols and benchmarks for video-language understanding. Inspir

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs

DGX agent

arXiv:2601.02023v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) increasingly utilize massive context windows as working memory for autonomous tasks, their reliability fluctua

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

nuTruck: Benchmarking Autonomous Driving Planning for Distributed Electric-drive Trucks

DGX agent

arXiv:2607.13704v1 Announce Type: new Abstract: The dominance of traditional rule-based methods in autonomous driving has gradually been replaced by learning-based approaches. While learning-based pla

model-releasesarxiv-cs-ro
16 Jul 2026
Model Releases

OccTrack360: 4D Panoptic Occupancy Tracking from Surround-View Fisheye Cameras

DGX agent

arXiv:2603.08521v2 Announce Type: replace Abstract: Understanding dynamic 3D environments in a spatially continuous and temporally consistent manner is fundamental for robotics and autonomous driving.

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

OvisOCR2 Technical Report

DGX agent

arXiv:2607.13639v1 Announce Type: cross Abstract: We introduce OvisOCR2, a 0.8B document parsing model. OvisOCR2 is designed as an end-to-end parser: given a document page image, it generates a Markdo

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Parsimonious disturbance-aware minimum-time planning with parametric uncertainty

DGX agent

arXiv:2607.13312v1 Announce Type: new Abstract: This study presents and validates a minimum-lap-time planning (MLTP) framework for motorsport applications that embeds robustness against both state dis

model-releasesarxiv-cs-ro
16 Jul 2026
Model Releases

Partially Correlated Verifier Cascades in LLM Harnesses: Concave Log-Odds, Polynomial Reliability, and Blind-Spot Ceilings

DGX agent

arXiv:2607.13918v1 Announce Type: cross Abstract: Serial verification gates are a core reliability primitive in LLM harnesses: a candidate answer is returned only if k verifier calls all accept it. Un

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Peak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic Assessment

DGX agent

arXiv:2607.13941v1 Announce Type: new Abstract: Video aesthetic assessment (VAA) aims to predict how aesthetically pleasing a video is, yet remains far less explored than other visual assessment tasks

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

PiVoT: A Variational Solution for Real-time Large-scale Multi-object Detection and Tracking under Heavy Clutter

DGX agent

arXiv:2607.13891v1 Announce Type: cross Abstract: Multi-object detection and tracking from noisy point clouds remain challenging in many data-scarce radar applications. Current Bayesian trackers based

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Policy of Thoughts: Scaling Test-Time Training for LLM Reasoning via Online Policy Evolution

DGX agent

arXiv:2601.20379v2 Announce Type: replace Abstract: Large language models (LLMs) struggle with complex, long-horizon reasoning due to instability caused by their frozen policy assumption. Current test

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

PQFA: Parallel Quantum Feature Augmentation of Fused Representations for Multimodal Classification

DGX agent

arXiv:2607.13466v1 Announce Type: new Abstract: Most multimodal learning methods improve how heterogeneous representations are aligned and fused, while post-fusion enhancement remains less explored. W

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

Quantum Circuit Vision: Cost-Aware Evaluation of Visual AI Agents for Quantum Code Generation

DGX agent

arXiv:2607.10057v1 Announce Type: cross Abstract: Can AI agents visually comprehend quantum circuit diagrams and generate verified executable code--and at what cost? We present Quantum Circuit Vision,

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

RAGthoven at SemEval-2026 Task 1: A Multi-Stage Pipeline Walks Into a Benchmark and Barely Clears the Bar

DGX agent

arXiv:2607.13189v1 Announce Type: cross Abstract: We present RAGthoven, our system for SemEval-2026 Task 1 (MWAHAHA), Subtask A (multilingual constrained humor generation in English, Spanish, and Chin

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Representation-Based Exploration for Language Models: From Test-Time to Post-Training

DGX agent

arXiv:2510.11686v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) promises to expand the capabilities of language models, but it is unclear if current RL techniques promote the dis

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

RF Spectrogram Anomaly Detection with Quantum Kitchen Sinks: Architecture, Representation, and Hardware Validation

DGX agent

arXiv:2607.13897v1 Announce Type: new Abstract: The broadcast nature of wireless channels exposes radio-frequency (RF) networks to anomalous and malicious transmissions, making anomaly detection a fun

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

S-squared-VLA: Decoupling Semantic and Spatial Streams in Vision-Language-Action Models for Autonomous Driving

DGX agent

arXiv:2607.13926v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated remarkable potential for high-level reasoning in autonomous driving, yet they fundamentally struggle to

model-releasesarxiv-cs-ro
16 Jul 2026
Model Releases

Safe Overtaking for Autonomous Racing Using Hierarchical Optimization and Learning-Based Control

DGX agent

arXiv:2607.13348v1 Announce Type: new Abstract: Autonomous racing overtaking requires balancing competitive performance with safety under nonlinear vehicle dynamics and real-time constraints. Model Pr

model-releasesarxiv-cs-ro
16 Jul 2026
Model Releases

Safeguard-Conditioned Uplift: Measuring Utility-Risk Frontiers for Dual-Use Biology Assistants

DGX agent

arXiv:2607.13039v1 Announce Type: cross Abstract: Safety evaluations for dual-use biology assistants often measure base-model capability, refusal behavior, or jailbreak success. These metrics miss a d

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

ScanFocus: A Coarse-to-Fine Framework for Spatio-Temporal Video Grounding

DGX agent

arXiv:2607.13421v1 Announce Type: cross Abstract: Spatio-Temporal Video Grounding (STVG) aims to retrieve the visual trajectory of a specific object from a video stream as described by a natural langu

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Securing LLMs in the Wild: Privacy and Security Challenges at the Edge

DGX agent

arXiv:2607.13088v1 Announce Type: cross Abstract: Large Language Models (LLMs) are rapidly moving from research settings into the wild, deployed on enterprise infrastructure, personal devices, and edg

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

Self-Supervised Visual Representation Learning: Pretrain-Finetuning or Joint Training?

DGX agent

arXiv:2607.13192v1 Announce Type: new Abstract: Self-supervision is a powerful technique for learning visual representations from unlabeled data. Existing techniques primarily adopt a two-stage approa

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Set-shifting Behavioral Test for Harnessed Agents

DGX agent

arXiv:2607.13396v1 Announce Type: new Abstract: What happens to an LLM agent's tool choice when the reliable tool silently changes within an ongoing session? We borrow set-shifting from cognitive psyc

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

SingGuard-NSFA: Extensible Guardrails for Agentic AI via Generative Reasoning and Real-Time Classification

DGX agent

arXiv:2607.13081v1 Announce Type: cross Abstract: We present nsfaguard, a guardrail framework for securing agentic AI systems against operational threats, such as prompt injection, sensitive informati

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

SPINE: Bridging the Cyber-Physical Gap with Agentic AI

DGX agent

arXiv:2607.13049v1 Announce Type: new Abstract: Foundation models have given robots a sophisticated brain for complex decision-making, yet deploying that intelligence into a physical platform still de

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

STOCKTAKE: Measuring the Gap Between Perception and Action in LLM Agents with a Fair Oracle

DGX agent

arXiv:2607.13618v1 Announce Type: new Abstract: LLM agents are increasingly evaluated on multi-week decision tasks in which the state that drives cost is never directly observed. On such tasks the fin

model-releasesarxiv-cs-ai
16 Jul 2026
← Previous
1…7071727374…361
Next →