AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

SoftVTBench: A Safety-Aware Visuo-Tactile Benchmark for Physically Constrained Robotic Manipulation of Deformable Objects

DGX agent

arXiv:2607.04234v1 Announce Type: cross Abstract: Deformable object manipulation poses challenges beyond task completion: successful execution must also maintain safe physical interaction, holding the

model-releasesarxiv-cs-ai
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Solve the Missing First Step: Can VLMs Standardize Raw Heterogeneous Medical Data?

DGX agent

arXiv:2607.04694v1 Announce Type: new Abstract: As vision-language models (VLMs) are increasingly applied to medical AI, existing benchmarks mainly focus on evaluating their diagnosis ability over giv

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

SovereignNegotiation-Bench: Evaluating User-Owned Personal Agents In Delegated Bargaining Under Privacy, Consent, Evidence, And Institutional Pressure

DGX agent

arXiv:2607.02814v1 Announce Type: cross Abstract: Personal agents will increasingly negotiate on behalf of users: splitting costs with other personal agents, appealing platform decisions, escalating s

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

SovereignPA-Bench: Evaluating User-Owned Personal Agents under Evolving Intent, Platform Mediation, and Consent Constraints

DGX agent

arXiv:2607.05363v1 Announce Type: new Abstract: Personal agents are becoming persistent user-owned intermediaries: they remember preferences, filter platform-mediated information, use tools, and negot

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

DGX agent

arXiv:2511.07403v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved remarkable progress in vision-language tasks, but continue to struggle with spatial rea

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models

DGX agent

arXiv:2607.05365v1 Announce Type: cross Abstract: Streaming speech-to-speech language models aim to answer spoken queries directly with synthetic speech. However, standard speech and text benchmarks d

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning

DGX agent

arXiv:2603.23483v2 Announce Type: replace-cross Abstract: Agentic multimodal large language models (MLLMs) (e.g., OpenAI o3 and Gemini Agentic Vision) achieve remarkable reasoning capabilities through

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Specific Domain Ontology Construction Using Large Language Models

DGX agent

arXiv:2606.20691v1 Announce Type: cross Abstract: Ontologies are useful structures to organize and maintain information that can be understood both by humans and systems. However, since their manual c

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Spectral Rewiring for Exploration, Purification, and Model Merging

DGX agent

arXiv:2607.03065v1 Announce Type: cross Abstract: Reinforcement learning has become a standard post-training recipe for large language models, but dense full-parameter updates create two deployment-re

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Spectral Signatures of Large Language Models

DGX agent

arXiv:2607.03377v1 Announce Type: cross Abstract: The rapidly growing repository of publicly available large language models (LLMs) presents significant challenges for systematic management and quanti

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Stacked LoRA for Subject-Adaptive EEG Foundation Models in Motor Imagery Decoding

DGX agent

arXiv:2607.03094v1 Announce Type: new Abstract: Electroencephalography (EEG) decoding for brain-computer interfaces (BCIs) faces a major challenge: substantial inter-subject variability limits effecti

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

SteelBench: Evaluating Vision-Language Models in Real-World Industrial Environments

DGX agent

arXiv:2607.05264v1 Announce Type: new Abstract: Existing video benchmarks evaluate action recognition on consumer videos, egocentric recordings, or simulated industrial environments. They do not test

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

STELLA: Efficient Sensor-to-LLM Translation for On-Device Human Activity Recognition

DGX agent

arXiv:2607.03089v1 Announce Type: cross Abstract: HAR is increasingly expected to run continuously on edge devices, yet recent LLM-based methods remain hard to deploy: raw sensor prompts are long, clo

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Streaming Model Cascades for Semantic SQL

DGX agent

arXiv:2604.00660v2 Announce Type: replace-cross Abstract: Modern data warehouses extend SQL with semantic operators that invoke large language models on each qualifying row, making per-row inference o

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Structured Prompting and Automated Evaluation in Fixed Synthetic Japanese-Language Counseling Dialogues

DGX agent

arXiv:2507.02950v3 Announce Type: replace-cross Abstract: Large language models (LLMs) may support counseling training, yet evidence from Japanese-language interactions and automated quality ratings r

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

StructuredEdit: Constraint-Aware Graphic Design Editing via Differentiable Parameter Propagation

DGX agent

arXiv:2607.04612v1 Announce Type: cross Abstract: Graphic design editing requires precise manipulation of typography, layout, and visual hierarchy under strict design constraints. Following the introd

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

SVG-EAR: Parameter-Free Linear Compensation for Sparse Video Generation via Error-aware Routing

DGX agent

arXiv:2603.08982v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) have become a leading backbone for video generation, yet their quadratic attention cost remains a major bottleneck. Sp

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

TabQueryBench: A Query-Centric Benchmark for Synthetic Tabular Data

DGX agent

arXiv:2607.03926v1 Announce Type: cross Abstract: Synthetic tabular data support use cases like data sharing, model development under access restrictions, and rapid prototyping of analytical workflows

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

TACG: Trajectory-Aware Commit Gating for Diffusion Language Model Decoding

DGX agent

arXiv:2607.03236v1 Announce Type: new Abstract: Diffusion language models (DLLMs) generate text by iteratively denoising masked positions, exposing a trajectory of predictive distributions rather than

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Taming I2V models for Image HOI Editing: A Cognitive Benchmark and Agentic Self-Correcting Framework

DGX agent

arXiv:2606.19073v2 Announce Type: replace Abstract: Current image editing methods excel at static attributes but fail at complex Human-Object Interactions (HOI), a critical challenge unaddressed by ex

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Target-Guided Selective Reweighting for Physics-Informed Neural Network Inverse Problems: A Transfer Learning Approach

DGX agent

arXiv:2607.05271v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) encounter ill-posed optimization, loss competition, and parameter compensation in partial differential equation

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Taste-aware music retrieval from audio embeddings

DGX agent

arXiv:2607.03296v1 Announce Type: cross Abstract: Crossmodal correspondences between sound and taste are well established in psychology and neuroscience, but largely absent from content-based multimed

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Teacher Supervision over Representation Equivalence Classes

DGX agent

arXiv:2607.03572v1 Announce Type: cross Abstract: Knowledge distillation is usually framed as a choice of what to match in the teacher - its logits, hidden features, or sample relations - which presup

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Teaching Code LLMs to Reason with Intermediate Formal Specifications

DGX agent

arXiv:2607.04232v1 Announce Type: cross Abstract: Unlike natural-language specifications, executable formal specifications provide machine-checkable constraints for verifying, debugging, and repairing

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

TESSERA v2: Scaling Pixel-wise Earth Foundation Models

DGX agent

arXiv:2607.03949v1 Announce Type: new Abstract: Pixel-wise Earth-observation (EO) foundation models are now achieving state-of-the-art performance via generated spatial embeddings. However, how these

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

TestMate: Test-Time Domain Adaptation Aided by Lightweight Vision Foundation Model

DGX agent

arXiv:2607.03810v1 Announce Type: new Abstract: Test-Time Domain Adaptation (TTDA) aims to adapt Deep Neural Networks to distribution shifts using only streaming, unlabeled test data in real time. Cur

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

TexTailor: Inference-Time Textual Guidance Tailoring for Multimodal Diffusion Transformers

DGX agent

arXiv:2601.02211v2 Announce Type: replace Abstract: Recent breakthroughs of transformer-based diffusion models, particularly with Multimodal Diffusion Transformers (MMDiT) driven models like FLUX and

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

The Good, the Bad, and the Brittle: Benchmarking Robustness and Generalisation of Histopathology Foundation Models

DGX agent

arXiv:2607.04401v1 Announce Type: new Abstract: How robust and generalisable are pathology foundation models and have their scaling limites been reached? We benchmarked twelve pathology foundation mod

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

The Map Behind the Flow: Finite-Step Gradient Descent as a Dynamical System

DGX agent

arXiv:2607.04993v1 Announce Type: cross Abstract: Many phenomena of deep learning are dynamical: they concern not only which minima exist, but how gradient descent reaches, avoids, or selects among th

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

The Method of Gaps: Exact Expressions for the Generalization Error of Supervised Learning Algorithms

DGX agent

arXiv:2411.12030v3 Announce Type: replace Abstract: In this paper, the method of gaps, a technique for deriving closed-form expressions in terms of information measures for the generalization error of

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

The Moving Target: A Longitudinal Audit of Trustworthiness Drift Across Twelve Checkpoints of Open-Source Chat LLMs

DGX agent

arXiv:2607.02587v1 Announce Type: cross Abstract: Model cards quote trust-benchmark scores without recording when they were measured, and the same number is routinely carried across successive checkpo

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

The Multipath Blind Spot: K-Agnostic Robust Calibration for Sparse-Anchor Metric Depth from Frozen Foundations

DGX agent

arXiv:2607.04101v1 Announce Type: new Abstract: Monocular depth foundations predict domain-general relative depth but lack absolute scale; a handful of sparse metric anchors from a range sensor can ca

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

The P^3 Dataset: Pixels, Points and Polygons for Multimodal Building Vectorization

DGX agent

arXiv:2505.15379v2 Announce Type: replace Abstract: We present the P^3 dataset, a large-scale multimodal benchmark for building vectorization, constructed from aerial LiDAR point clouds, high-resoluti

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

The Remarkable Effectiveness of Providing AI Agents with Natural Language Tools: A Replication Study Validating NLT Performance Across 14 Models

DGX agent

arXiv:2607.03953v1 Announce Type: cross Abstract: This study independently replicates and extends the Natural Language Tools (NLT) framework of Johnson et al.~(2025), which questions the use of struct

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

The Role of Prompt Language and Translation-Theory-Driven Prompts in Large Language Models: A Case Study on Spanish-Chinese Journalistic Translation

DGX agent

arXiv:2607.03160v1 Announce Type: cross Abstract: This study examines how prompt language and translation theory-driven prompt design influence the quality of Spanish-Chinese journalistic translations

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

The Truncation Blind Spot: How Decoding Strategies Systematically Exclude Human-Like Token Choices

DGX agent

arXiv:2603.18482v2 Announce Type: replace Abstract: Standard decoding strategies for text generation, including top-k, nucleus sampling, and contrastive search, select tokens based on likelihood, rest

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Think Deep, Not Just Long: Measuring LLM Reasoning Effort via Deep-Thinking Tokens

DGX agent

arXiv:2602.13517v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive reasoning capabilities by scaling test-time compute via long Chain-of-Thought (CoT). Howev

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Tile-Level Activation Overlap for Efficient LLM Inference

DGX agent

arXiv:2607.02521v1 Announce Type: cross Abstract: SwiGLU is the dominant MLP activation in modern large language models, yet its intermediate tensor materialization costs 9-37% of MLP execution time.

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

TiROD: Tiny Robotics Dataset and Benchmark for Continual Object Detection

DGX agent

arXiv:2409.16215v4 Announce Type: replace-cross Abstract: Detecting objects with visual sensors is crucial for numerous mobile robotics applications, from autonomous navigation to inspection. However,

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

TokSuite: Measuring the Impact of Tokenizer Choice on Language Model Behavior

DGX agent

arXiv:2512.20757v2 Announce Type: replace Abstract: Tokenizers provide the fundamental basis through which text is represented and processed by language models (LMs). Despite the importance of tokeniz

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

ToolFailBench: Diagnosing Tool-Use Failures in LLM Agents

DGX agent

arXiv:2607.04686v1 Announce Type: cross Abstract: Tool calling is central to modern language model agents, but aggregate benchmark scores often hide where tool use fails. A model that never calls a ne

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Topology-Driven Transferability Estimation for 3D Medical Vision Foundation Models

DGX agent

arXiv:2607.04199v1 Announce Type: new Abstract: The growing number of medical vision foundation models highlights the need for effective model selection. However, mainstream selection methods rely on

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Toward Efficient Agents: Memory, Tool learning, and Planning

DGX agent

arXiv:2601.14192v2 Announce Type: replace Abstract: Recent years have witnessed increasing interest in extending large language models into agentic systems. While the effectiveness of agents has conti

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Toward Trustworthy Large Language Model Agents in Healthcare

DGX agent

arXiv:2607.05055v1 Announce Type: new Abstract: Healthcare appointment scheduling remains a persistent operational bottleneck, driven by manual coordination, fragmented legacy systems, and high admini

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Towards Open-World Referring Expression Comprehension: A Benchmark with Training-free Multi-task Consistency Checker

DGX agent

arXiv:2605.25706v2 Announce Type: replace Abstract: Referring expression comprehension (REC) aims to localize a target object within an image based on a given expression. Although recent advances in v

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Towards Realistic Remote Sensing Dataset Distillation with Discriminative Prototype-guided Diffusion

DGX agent

arXiv:2601.15829v2 Announce Type: replace Abstract: Recent years have witnessed the remarkable success of deep learning in remote sensing image interpretation, driven by the availability of large-scal

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Towards Reliable Local Security Agents: Verifiable Post-Training for Linux Privilege Escalation

DGX agent

arXiv:2603.17673v2 Announce Type: replace-cross Abstract: LLM agents are becoming increasingly important in the security domain, but leading systems are often closed-source, cloud-based, hard to repro

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Towards Standardized Light Field Quality Assessment: Hybrid Subjective Benchmarking and Objective Metric Evaluation

DGX agent

arXiv:2607.03494v1 Announce Type: new Abstract: Benchmarking immersive media coding solutions, especially in the standardization context, requires reliable and reproducible subjective quality assessme

model-releasesarxiv-cs-cv
7 Jul 2026
← Previous
1…9091929394…361
Next →