AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

TMI: Text-to-Image Meets Image-to-Image for Complementary Data Synthesis to Boost Long-Tailed Instance Segmentation

DGX agent

arXiv:2607.08201v1 Announce Type: cross Abstract: Large-vocabulary instance segmentation is constrained by long-tailed category distributions and fine-grained inter-class ambiguity. While data synthes

model-releasesarxiv-cs-ai
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TOPO-Bench: An Open-Source Topological Mapping Evaluation Framework with Quantifiable Perceptual Aliasing

DGX agent

arXiv:2510.04100v2 Announce Type: replace-cross Abstract: Topological mapping offers a compact and robust representation for navigation, but progress in the field is hindered by the lack of standardiz

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Towards Precision Therapy in Hepatocellular Carcinoma: A Clinical-Reasoning LLM for Risk Stratification and Treatment Guidance

DGX agent

arXiv:2607.08602v1 Announce Type: new Abstract: Hepatocellular carcinoma (HCC) is a common malignancy and a leading cause of cancer-related mortality. Current guidelines and staging systems provide co

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Training and Evaluating Diffusion Policies with Long Context Lengths

DGX agent

arXiv:2606.16447v2 Announce Type: replace-cross Abstract: Imitation learning has enabled highly-dexterous robotic manipulation from RGB observations. Policies trained with these methods, however, typi

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

TVTA: Trajectory-Aware Viseme-Guided Temporal Aggregation for Event-Based Lip Reading

DGX agent

arXiv:2607.08236v1 Announce Type: new Abstract: Event-based lip reading has recently emerged as a promising direction for visual speech recognition, benefiting from the high temporal resolution and mo

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

UAV-OVVIS: Unmanned Aerial Vehicles Also Need Open-Vocabulary Video Instance Segmentation

DGX agent

arXiv:2607.08075v1 Announce Type: new Abstract: Unmanned Aerial Vehicle (UAV) videos are widely used in traffic monitoring, urban management, and emergency rescue. However, existing UAV video percepti

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Uncertainty-gated selection for block-sparse attention

DGX agent

arXiv:2607.07724v1 Announce Type: cross Abstract: Block-sparse attention scales long-context language models by replacing the O(N^2) softmax with a per-query top-k selection over key blocks. This cuto

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio

DGX agent

arXiv:2607.08127v1 Announce Type: new Abstract: Generative video foundation models exhibit strong compositional priors, yet world-action models (WAMs) and video-action models (VAMs) often lose these p

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Understanding Axes of Difficulty For Long Context Tasks Via PredicateLongBench

DGX agent

arXiv:2607.08284v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated rapidly improving long-context capabilities, prompting a wave of benchmarks designed to evaluate them. Ho

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks

DGX agent

arXiv:2607.08768v1 Announce Type: new Abstract: The rapid development of large language models and multimodal large language models has accelerated the emergence of proactive agents capable of operati

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

UniRef-UAV: A Multimodal Benchmark for Universal Referring in UAV Imagery

DGX agent

arXiv:2607.08267v1 Announce Type: new Abstract: Unmanned aerial vehicles (UAVs) increasingly rely on visual grounding capabilities to localize task-relevant targets from diverse instructions in comple

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Validity of LLMs as data annotators: AMALIA on authority

DGX agent

arXiv:2607.08731v1 Announce Type: cross Abstract: A national language model offers a linguistic community its own instrument for measuring what its citizens say and value. Portugal's AMALIA, a publicl

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Variational Phasor Circuits for Phase-Native Brain-Computer Interface Classification

DGX agent

arXiv:2603.18078v2 Announce Type: replace Abstract: We present the Variational Phasor Circuit (VPC), a deterministic classical learning architecture on the continuous S^1 unit-circle manifold. Inspire

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

VSRo-200: A Romanian Visual Speech Recognition Dataset for Studying Supervision and Multimodal Robustness

DGX agent

arXiv:2607.08112v1 Announce Type: new Abstract: We introduce VSRo-200, the first large-scale dataset for visual speech recognition (lip reading) in Romanian, comprising 200 hours of real-world podcast

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

WaspMOT: A Benchmark for Long-Term Multi-Object Tracking of Trichogramma Wasps

DGX agent

arXiv:2607.08729v1 Announce Type: new Abstract: Multi-object tracking (MOT) has achieved strong performance on benchmarks dominated by short video sequences. However, such datasets do not adequately e

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving

DGX agent

arXiv:2607.08375v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving. However, existing methods either lack comprehensive world cognition o

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents

DGX agent

arXiv:2607.08032v1 Announce Type: new Abstract: Large language models, and the agents built on them, spend an ever-growing share of their compute and memory on remembering: caching attention keys and

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

When LLMs Agree, Are They Right? Auditing Self-Consistency and Cross-Model Agreement as Confidence Signals

DGX agent

arXiv:2607.08065v1 Announce Type: new Abstract: LLM-as-judge (Zheng et al., 2023) is increasingly the default for evaluating AI systems in enterprise pipelines, often scaled to ensembles (Verga et al.

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

When the Judge Changes, So Does the Measurement: Auditing LLM-as-Judge Reliability

DGX agent

arXiv:2607.08535v1 Announce Type: cross Abstract: An LLM-as-judge score can move even when the candidate responses stay fixed, simply because the evaluator has changed. We treat this evaluator-replace

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

When Thinking Hurts: Epistemic Signals in the Reasoning Chains of Visual Language Models

DGX agent

arXiv:2607.08059v1 Announce Type: cross Abstract: Uncertainty quantification for visual language models (VLMs) conventionally targets the answer token distribution. We provide the first three-family e

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Who Analyses the Analyser? Self-Validating LLM Hazard Analysis with Constitutional Meta-STPA

DGX agent

arXiv:2607.08054v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly trusted to draft the artifacts of safety analysis such as, losses, hazards, Unsafe Control Actions (UCAs

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Workload-Preserving Differentially Private Synthetic Data for Causal Inference via Maximum-Entropy Calibration

DGX agent

arXiv:2607.08122v1 Announce Type: new Abstract: Workload-based differentially private (DP) synthetic data methods privately measure aggregate queries and post-process the noisy answers into synthetic

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

XOV-Action: Towards Generalizable Open-Vocabulary Action Recognition

DGX agent

arXiv:2403.01560v3 Announce Type: replace Abstract: Inspired by the impressive success of image-text foundation models, recent works have proposed to adapt these foundation models to video data, leadi

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

A Gold-Standard Study of What Makes a Lightweight Game-Playing Agent Strong

DGX agent

arXiv:2607.06854v1 Announce Type: cross Abstract: Reinforcement learning agents for imperfect-information card games are only as strong as the opponents they train against, and they are hard to grade,

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

A knowledge-augmented dataset of high-risk driving scenarios with LLM annotations for autonomous driving

DGX agent

arXiv:2607.07103v1 Announce Type: new Abstract: Safe autonomous driving requires both rapid responses to common high-risk events and deeper reasoning over rare, extreme long-tail scenarios in traffic

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

A Study of Commonsense Reasoning over Visual Object Properties

DGX agent

arXiv:2508.10956v3 Announce Type: replace-cross Abstract: Inspired by human categorization, visual reasoning about object properties, such as physical attributes and functions, involves identifying an

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Ace! Motion Planning of Professional-Level Table Tennis Serves with a Robot Arm

DGX agent

arXiv:2607.06989v1 Announce Type: new Abstract: Table tennis, a dynamic, compact, and popular sport, has received significant attention as a robotics benchmark over the last decades. Most of the resea

model-releasesarxiv-cs-ro
9 Jul 2026
Model Releases

AgentLens: Production-Assessed Trajectory Reviews for Coding Agent Evaluation

DGX agent

arXiv:2607.06624v1 Announce Type: new Abstract: We present AgentLens, a production-assessed benchmark for interactive code agents. Most code-agent benchmarks reduce a run to a single bit -- did the ta

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning

DGX agent

arXiv:2607.07690v1 Announce Type: cross Abstract: Reinforcement learning from verifiable rewards (e.g. GRPO) is the engine behind today's reasoning models, yet it grades only the final answer. On hard

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

AI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluation

DGX agent

arXiv:2602.05088v4 Announce Type: replace Abstract: Millions of people now use generative AI chatbots for psychological support. Despite their promise, the most pressing question in AI for mental heal

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

An optimal control approach for neural network architecture adaptation with a posteriori error estimation

DGX agent

arXiv:2607.07637v1 Announce Type: new Abstract: This work presents a novel approach for adapting neural network architecture along the depth based on a posteriori error estimation. By formulating neur

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

ASFR-Net: Adversarial Alignment and Spatio-Frequency Refinement Network for Heterogeneous Remote Sensing Image Change Detection

DGX agent

arXiv:2607.07161v1 Announce Type: new Abstract: The core challenge of heterogeneous change detection in remote sensing imagery lies in effectively decoupling genuine land-cover changes from significan

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

At-Grok Is Not Converged:A Measurement-Validity Audit for Grokking Representation Metrics

DGX agent

arXiv:2607.06639v1 Announce Type: cross Abstract: On modular arithmetic, a network's embedding keeps compressing for tens of thousands of steps after it has already generalized. Reading effective rank

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Beyond Attack-Success Rate: Action-Graded Severity Scale for Tool-Using AI Agents

DGX agent

arXiv:2607.07474v1 Announce Type: cross Abstract: Agentic red-teaming benchmarks report whether an injected agent was compromised as a single bit: the attack succeeded, or it did not. We argue that th

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Bifidelity Parameter Estimation Using Conditional Diffusion Models

DGX agent

arXiv:2504.01894v2 Announce Type: replace Abstract: We present a bifidelity method for uncertainty quantification of parameter estimates in complex systems, leveraging generative models trained to sam

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

Breaking Database Lock-in: Agentic Regeneration of High Performance Storage Readers for Database Bypass

DGX agent

arXiv:2607.07696v1 Announce Type: cross Abstract: Analytical workloads operating on data stored in external database systems face a fundamental bottleneck: data access is guarded entirely by the datab

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

BubbleSH: A Dataset of Rising Bubbles with Deformable Interfaces

DGX agent

arXiv:2607.07275v1 Announce Type: new Abstract: Bubbly flows exhibit complex multiscale dynamics, with deformable bubbles interacting through the surrounding liquid and giving rise to strongly coupled

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages

DGX agent

arXiv:2607.06596v1 Announce Type: cross Abstract: Trusted monitoring is a central defense in AI control: a cheaper trusted model scores an untrusted model's actions for sabotage, and the most suspicio

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

CaLiSym: Learning Symplectic Dynamics of Real-World Systems through Structured Canonical Lifts

DGX agent

arXiv:2607.06824v1 Announce Type: cross Abstract: Physics-informed learning promises data-efficient and stable dynamics prediction, yet its strongest geometric guarantees have largely remained confine

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

Can Reinforcement Learning Efficiently Discover Price Manipulation?

DGX agent

arXiv:2607.06121v1 Announce Type: cross Abstract: In this paper, we investigate whether a model-free RL agent can identify and exploit price manipulation opportunities more effectively than a traditio

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Co-LMLM: Continuous-Query Limited Memory Language Models

DGX agent

arXiv:2607.07707v1 Announce Type: cross Abstract: Limited memory language models (LMLMs) externalize factual knowledge during pretraining to a knowledge base (KB), rather than memorizing it in their w

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Communicative Efficiency of Single vs. Multi-Axis Robot Neck Motion

DGX agent

arXiv:2607.07390v1 Announce Type: new Abstract: Nonverbal communication through head and neck movement is fundamental to human social signalling, yet how robotic neck morphology translates motion into

model-releasesarxiv-cs-ro
9 Jul 2026
Model Releases

Comparative Study of Domain-adapted VLMs for General Document Visual Question Answering

DGX agent

arXiv:2607.07179v1 Announce Type: new Abstract: Document Visual Question Answering (DocVQA) presents a complex multimodal challenge, requiring models to exploit visual, textual, and layout information

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

CompDiff: Hierarchical Compositional Diffusion for Fair and Zero-Shot Intersectional Medical Image Generation

DGX agent

arXiv:2603.16551v2 Announce Type: replace-cross Abstract: Generative models are increasingly used to augment medical imaging datasets for fairer AI, yet a key assumption often goes unexamined: that ge

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Context-Aware Force Estimation for Deformable Tool Manipulation in Robotic Environmental Swabbing via Few-Shot Continual Adaptation

DGX agent

arXiv:2607.07574v1 Announce Type: new Abstract: Robotic surface swabbing requires sustained interaction between a compliant tool and heterogeneous environments, where accurate estimation of tip-level

model-releasesarxiv-cs-ro
9 Jul 2026
Model Releases

Cost-Effective Agent Harnesses for Abstract Reasoning and Generalization on ARC-AGI-1

DGX agent

arXiv:2607.06764v1 Announce Type: new Abstract: Recent progress on ARC-AGI-1 from disclosed architectures has come broadly from two regimes: heavy test-time compute over frontier models (evolutionary

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Diffusion Models in Simulation-Based Inference: A Tutorial Review

DGX agent

arXiv:2512.20685v3 Announce Type: replace-cross Abstract: Diffusion models have recently emerged as powerful learners for simulation-based inference (SBI), enabling fast and accurate estimation of lat

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

Does AI Understand Imaging? A Systematic Benchmark of Agentic AI for Computational Imaging Tasks

DGX agent

arXiv:2607.07189v1 Announce Type: new Abstract: Vision-language models (VLMs) and agentic AI have shown strong performance on semantic visual tasks, but it remains unclear whether they can handle the

model-releasesarxiv-cs-ai
9 Jul 2026
← Previous
1…7778798081…361
Next →