AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

TQA-Bench: Evaluating LLMs for Multi-Table Question Answering

DGX agent

arXiv:2411.19504v2 Announce Type: replace Abstract: The advance of large language models (LLMs) has unlocked great opportunities in complex multi-modal data management tasks, particularly in question

model-releasesarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Training-Free Generalized Few-Shot Segmentation through Open-Vocabulary Semantic Arbitration

DGX agent

arXiv:2606.09474v1 Announce Type: new Abstract: Generalized Few-Shot Semantic Segmentation (GFSS) has traditionally been approached as a representation-learning problem, requiring task-specific adapta

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Trajectory-Refined Distillation

DGX agent

arXiv:2606.08432v1 Announce Type: new Abstract: On-policy distillation (OPD) has become a central post-training tool for large language models (LLMs), providing dense per-token teacher supervision alo

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TriHead-GAN: A Generative Adversarial Network with Triple-Head Discriminator for Carbon Emission Time Series Generation

DGX agent

arXiv:2606.07569v1 Announce Type: new Abstract: Accurate carbon emission monitoring is critical for climate policy and emerging regulatory mechanisms such as the EU Carbon Border Adjustment Mechanism,

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

TRL-Bench: Standardizing Cross-Paradigm Representation-Level Evaluation of Tabular Encoders

DGX agent

arXiv:2606.09323v1 Announce Type: new Abstract: Tabular encoders are usually evaluated inside task-specific end-to-end pipelines, so models from different training paradigms are difficult to compare d

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TT-DAC-PS: Twin-Target Deterministic Actor-Critic with Policy Smoothing for Optimal Trade Execution

DGX agent

arXiv:2606.08379v1 Announce Type: new Abstract: This study addresses the optimal execution of large stock sell programs by introducing TT-DAC-PS (Twin-Target Deterministic Actor-Critic with Policy Smo

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Understanding Benchmark Language Under Weakened Formal Semantics

DGX agent

arXiv:2509.17455v2 Announce Type: replace-cross Abstract: State-of-the-art NLP benchmarks require interpretation of natural language that specifies conditions, procedures, and exceptions, often relyin

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Understanding the Parameter Space Geometry of Transformers Encoding Boolean Functions

DGX agent

arXiv:2606.08768v1 Announce Type: new Abstract: Transformers consistently fail to learn certain simple functions that are provably expressible with specific parameter settings. This gap between learna

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Unification of Closed-Open Industrial Detection Scenarios: New Large-Scale Benchmarks,Challenges and Baselines

DGX agent

arXiv:2606.07953v1 Announce Type: new Abstract: Large-scale Visual-Language Models (LVLMs) have achieved remarkable success in natural visual tasks, yet their application to industrial defect detectio

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

UniQL: Towards Dialect-Universal Benchmarking for Text-to-SQL

DGX agent

arXiv:2606.08018v1 Announce Type: new Abstract: Existing text-to-SQL benchmarks are largely centered on SQLite, making it difficult to evaluate whether models can generalize across heterogeneous SQL d

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Unsupervised Partner Design Enables Robust Ad-hoc Teamwork

DGX agent

arXiv:2508.06336v2 Announce Type: replace-cross Abstract: We introduce Unsupervised Partner Design (UPD), a population-free multi-agent reinforcement learning method for robust ad-hoc teamwork. UPD ge

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

DGX agent

arXiv:2606.07992v1 Announce Type: new Abstract: As the Model Context Protocol (MCP) standardizes tool-calling for autonomous agents, it introduces a critical, unexamined attack surface: the error-hand

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

VideoWeaver: Evaluating and Evolving Skills for Agentic Long Video Generation

DGX agent

arXiv:2606.08091v1 Announce Type: new Abstract: Recent agent frameworks such as Claude Code, Codex, and OpenClaw are strong at tool use and orchestration, but whether they can handle long video genera

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Visual Template Inference for Data Extraction from Documents

DGX agent

arXiv:2501.06659v2 Announce Type: replace-cross Abstract: Many templatized documents are programmatically generated from structured data following a visual template. Such documents include invoices, t

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning?

DGX agent

arXiv:2606.07872v1 Announce Type: new Abstract: When a multimodal large language model answers a visual reasoning question correctly, is the prediction actually supported by the task-critical visual e

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Agents

DGX agent

arXiv:2606.07595v1 Announce Type: cross Abstract: Vision-language agents increasingly consume screenshots, documents, and user interfaces before writing to memory, sending messages, or invoking extern

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

vla.cpp: A Unified Inference Runtime for Vision-Language-Action Models

DGX agent

arXiv:2606.08094v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically shipped as Python/PyTorch stacks that assume a workstation-class GPU, a mismatch for the hardware

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation

DGX agent

arXiv:2606.07723v1 Announce Type: new Abstract: Open-vocabulary long-horizon manipulation requires robots to reason over flexible instructions and complex multi-object scenes while adaptively planning

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces

DGX agent

arXiv:2606.09426v1 Announce Type: new Abstract: Computer-use agents (CUAs) increasingly operate in runtimes that combine visual desktop control, command-line execution, code editing, browsers, and ext

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

When Behavioral Safety Evaluation Fails: A Representation-Level Perspective

DGX agent

arXiv:2606.08044v1 Announce Type: cross Abstract: Large Language Model (LLM) safety has often been evaluated at the behavior level, which provides limited evidence of internal robustness, as these eva

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

When Benign Inputs Lead to Severe Harms: Eliciting Unsafe Unintended Behaviors of Computer-Use Agents

DGX agent

arXiv:2602.08235v2 Announce Type: replace-cross Abstract: Although computer-use agents (CUAs) hold significant potential to automate increasingly complex OS workflows, they can demonstrate unsafe unin

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

When Do Local Score Models Extrapolate Across Size? A Diagnostic Theory and Benchmark

DGX agent

arXiv:2606.09705v1 Announce Type: new Abstract: Scientific generative modeling often requires size transfer, where models trained on small systems are evaluated on larger ones. While translation-invar

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Where Does the Answer Come From? Benchmarking View-Level Visual Evidence Identification in Multi-View MLLMs for Autonomous Driving

DGX agent

arXiv:2606.09644v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) achieve strong results on visual reasoning benchmarks, but answer accuracy alone does not indicate whether a

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Where Instruction Hierarchy Breaks: Diagnosing and Repairing Failures in Reasoning Language Models

DGX agent

arXiv:2606.07808v1 Announce Type: new Abstract: Reasoning language models deployed in agentic workflows must follow an instruction hierarchy: when instructions from different sources conflict, the mod

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Who Earns the Safety? Intervention-Aware Quantum Predictive Control with Safety Attribution

DGX agent

arXiv:2606.09778v1 Announce Type: cross Abstract: Hard safety filters are increasingly placed downstream of learned controllers to guarantee constraint satisfaction at run time. Yet a filtered control

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Your Model Already Knows: Attention-Guided Safety Filter for Vision-Language-Action Models

DGX agent

arXiv:2606.09749v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated impressive end-to-end performance across a variety of robotic manipulation tasks. However, these

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Zero-Parameter Geometric Gating for Temporally Stable Low-Altitude UAV Video Semantic Segmentation

DGX agent

arXiv:2606.09162v1 Announce Type: new Abstract: Video semantic segmentation for low-altitude UAVs requires temporal consistency, yet dense optical flow introduces spatially structured noise in the pla

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Zero-Shot Learning in Industrial Scenarios: New Large-Scale Benchmark, Challenges and Baseline

DGX agent

arXiv:2606.07965v1 Announce Type: new Abstract: Large Visual Language Models (LVLMs) have achieved remarkable success in vision tasks. However, the significant differences between industrial and natur

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Zero-Shot Semantic Re-Identification for Autonomous Driving: A VLM Baseline Study

DGX agent

arXiv:2606.09362v1 Announce Type: new Abstract: Re-Identification (ReID) in autonomous driving is typically formulated as a visual matching problem, where observations of vehicles, pedestrians, and cy

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

ZIPP:Zero-shot Image Personalization from Personas

DGX agent

arXiv:2606.08841v1 Announce Type: new Abstract: Text-to-image diffusion models are increasingly deployed in open-ended creative contexts, yet their outputs remain impersonal, optimized for aggregate a

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

3DMorph: Single-Image-Guided Local 3D Shape Editing and Morphing

DGX agent

arXiv:2606.07115v1 Announce Type: new Abstract: Despite recent progress in 3D generation, intuitive editing of existing shapes remains limited. Unlike images, which benefit from well-established inpai

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

A Comprehensive Anatomy of Human and DeepSeek-R1 LLM Mathematical Reasoning

DGX agent

arXiv:2606.07410v1 Announce Type: cross Abstract: The emergence of 'Aha moments' in large language models, particularly DeepSeek-R1-0120, has raised the question of whether these systems genuinely rea

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

A Cross-view Fusion Framework for Robust 6-DoF Grasp Pose Estimation

DGX agent

arXiv:2606.06878v1 Announce Type: cross Abstract: In this paper, we propose a cross-view fusion framework that enhances the robustness of 6-DoF grasp pose estimation in corner views. Our framework all

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

A Four-Condition Diagnostic Protocol for Evidence Utilization in Long-Context and Retrieval-Augmented Language Models

DGX agent

arXiv:2606.06758v1 Announce Type: new Abstract: Final-answer accuracy, retrieval recall, and citation overlap do not by themselves identify whether a long-context or retrieval-augmented language model

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

A Geometric Gaussian Mixture Representation of Plane Curves

DGX agent

arXiv:2606.06505v1 Announce Type: cross Abstract: We introduce a user defined probabilistic polygonal representation for plane curves. Given a curve, we select vertices on the curve and connect consec

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

A Held-Out Transition-Pair Falsifier for Long-Horizon Non-Abelian State Tracking

DGX agent

arXiv:2606.07254v1 Announce Type: new Abstract: State tracking exposes a sharp limitation of sequence models: the relevant signal is often not a summary of observed tokens, but an ordered latent state

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Accelerating Multi-Objective Bayesian Optimisation via Predictive-Gradient Catalysts

DGX agent

arXiv:2606.06984v1 Announce Type: new Abstract: This paper presents a general acceleration mechanism for multi-objective Bayesian optimisation (MOBO) that leverages Gaussian process predictive gradien

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Act As a Real Researcher: A Suite of Benchmarks Evaluating Frontier LLMs and Agentic Harnesses in Research Lifecycle

DGX agent

arXiv:2606.07462v1 Announce Type: new Abstract: As foundation models advance and agent scaffolding becomes increasingly sophisticated, agents have demonstrated remarkable proficiency in complex, long-

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

ActiveGrasp: Information-Guided Active Grasping with Calibrated Energy-based Model

DGX agent

arXiv:2511.12795v2 Announce Type: replace Abstract: Grasping in a densely cluttered environment is a challenging task for robots. Previous methods tried to solve this problem by actively gathering mul

model-releasesarxiv-cs-ro
8 Jun 2026
Model Releases

ADAGE: Active Defenses Against GNN Extraction

DGX agent

arXiv:2503.00065v4 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) achieve high performance in various real-world applications, such as drug discovery, traffic states prediction, a

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Adaptive Conditional Forest Sampling for Spectral Risk Optimisation under Decision-Dependent Uncertainty

DGX agent

arXiv:2603.12507v2 Announce Type: replace Abstract: Minimising a spectral risk objective, defined as a weighted combination of expected cost and Conditional Value-at-Risk (CVaR), is challenging when t

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

An Analysis Focused on Womens Safety: Can VAD Models Be Enhanced by a Multi-modal Dataset?

DGX agent

arXiv:2605.25806v2 Announce Type: replace Abstract: Women's safety and security are paramount for a modern society. Crimes against women occur in daylight as well as in low-light conditions. Often, su

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

An Integrated Roadside Sensing and Communication Framework for Vulnerable Road User Safety at Signalized Intersections

DGX agent

arXiv:2606.07016v1 Announce Type: cross Abstract: Vulnerable road users (VRUs) account for approximately half of urban traffic deaths globally, with intersections concentrating a disproportionate shar

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Architecture Shapes Transfer Specificity in Implicit Neural Representations

DGX agent

arXiv:2606.06827v1 Announce Type: new Abstract: Transfer in coordinate networks is often measured by warm-start gain, but whether that gain reflects source-specific structure or generic weight reuse i

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models

DGX agent

arXiv:2606.06534v1 Announce Type: cross Abstract: Longitudinal medical visual question answering (VQA) requires reasoning about anatomical differences between an image of a current time point and an i

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Audio-Visual World Models: Grounding Multisensory Imagination for Embodied Agents

DGX agent

arXiv:2512.00883v3 Announce Type: replace-cross Abstract: World models simulate environmental dynamics to enable agents to plan and reason about future states. While existing approaches have primarily

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Auditing Training Data in Domain-adapted LLMs: LoRA-MINT

DGX agent

arXiv:2606.06946v1 Announce Type: cross Abstract: We present LoRA-MINT, a new methodology for Membership Inference Test (MINT) applied to recent Large Language Models (LLMs) fine-tuned for specific Na

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Aumann-SHAP: The Geometry of Counterfactual Interaction Explanations in Machine Learning

DGX agent

arXiv:2603.14014v2 Announce Type: replace Abstract: We introduce Aumann-SHAP, an interaction-aware framework that decomposes counterfactual transitions by restricting the model to a local hypercube co

model-releasesarxiv-cs-lg
8 Jun 2026
← Previous
1…146147148149150…361
Next →