AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Revealing Training Data Exposure in Vision Language Large Models via Parameter Gradients

DGX agent

arXiv:2606.24774v1 Announce Type: new Abstract: Vision-Language Large Models (VLLMs) trained on massive crawled corpora raise pressing copyright and data-provenance concerns. These concerns are partic

model-releasesarxiv-cs-cv
24 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Reward-Centered ReST-MCTS: A Robust Decision-Making Framework for Robotic Manipulation in High Uncertainty Environments

DGX agent

arXiv:2503.05226v2 Announce Type: replace-cross Abstract: Monte Carlo tree search is attractive for robotic manipulation because it can improve action selection through simulation without requiring a

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

RoPE-Aware Bit Allocation for KV-Cache Quantization

DGX agent

arXiv:2606.24033v1 Announce Type: cross Abstract: Existing low-bit KV-cache quantizers often treat each cached key as a flat vector. Under RoPE, however, a key's contribution to a future attention log

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Rule2Text: A Framework for Generating and Evaluating Natural Language Explanations of Knowledge Graph Rules

DGX agent

arXiv:2508.10971v2 Announce Type: replace-cross Abstract: Knowledge graphs (KGs) can be enhanced through rule mining; however, the resulting logical rules are often difficult for humans to interpret d

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

SAFARI: Scaling Long Horizon Agentic Fault Attribution via Active Investigation

DGX agent

arXiv:2606.24626v1 Announce Type: new Abstract: As autonomous agents tackle increasingly complex multi-step, multi-agent tasks, their execution trajectories have scaled beyond the constraints of even

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

SciZoom: A Large-scale Benchmark for Hierarchical Scientific Summarization across the LLM Era

DGX agent

arXiv:2603.16131v2 Announce Type: replace Abstract: The explosive growth of AI research has created unprecedented information overload, increasing the demand for scientific summarization at multiple l

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

SEAGAN: domain-Specific and Edge-Aware Graph Attention Network for Dynamic Plant Processes

DGX agent

arXiv:2606.19623v2 Announce Type: replace Abstract: Graph neural networks (GNNs) offer a flexible framework for learning from scientific data with physical, biological, or functional associations. One

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Self-Recognition Finetuning can Prevent and Reverse Emergent Misalignment

DGX agent

arXiv:2606.23700v1 Announce Type: cross Abstract: Emergent misalignment (EM) has been linked to the activation of misaligned persona vectors and evil character traits, suggesting that EM operates thro

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

SER: Learning to Ground Video Reasoning with Semantic Evidence Rewards

DGX agent

arXiv:2606.24726v1 Announce Type: new Abstract: Video MLLMs often struggle with fine-grained spatio-temporal reasoning, sometimes generating correct answers based on irrelevant frames or objects. Alth

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

SICI: A Semantic-Pragmatic Complexity Index Reveals Regime Shifts in LLM Stance Detection

DGX agent

arXiv:2606.13189v2 Announce Type: replace Abstract: Prompt-based LLMs are increasingly used for stance detection, but harder examples are not always repaired by clearer instructions, reasoning prompts

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

SignNet-1M: Large-Scale Multilingual Sign Language Video Dataset with Downstream Benchmarks

DGX agent

arXiv:2606.24361v1 Announce Type: new Abstract: Sign language models are typically trained on datasets captured under constrained conditions, with limited viewpoint, background, and signer-identity di

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

SP-Mind: An Autonomous Reasoning Agent for Spatial Proteomics Analysis

DGX agent

arXiv:2606.24235v1 Announce Type: new Abstract: Spatial proteomics enables single-cell-resolution characterization of protein expression within tissue architecture, playing a critical role in understa

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Structural Kolmogorov-Arnold Convolutions: Learnable Function on the Values or the Filter Shape as Parameter-Efficient Alternative to Per-Edge Convolutional KANs

DGX agent

arXiv:2606.24371v1 Announce Type: cross Abstract: Convolutional Kolmogorov--Arnold Networks (KANs) replace the fixed weights of a convolutional kernel with learnable univariate functions. The dominant

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

SURGELLM: Rethinking Multi-Task Evaluation through Task-Aware Feature Gating with Class-Balanced Normalization

DGX agent

arXiv:2606.24259v1 Announce Type: cross Abstract: Fine-tuned encoders deployed across heterogeneous NLP tasks face three compounding problems: mismatched inductive biases, class-imbalance corruption o

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Synergizing Physically Constrained MCMC and Chemical-Informed Gaussian Processes for Reaction Network Discovery

DGX agent

arXiv:2606.23757v1 Announce Type: cross Abstract: Extracting interpretable governing equations from sparse, noisy chemical time-series data remains difficult because discrete reaction topology and con

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Systematic Exploration of 4-Expert Heterogeneous Mixture-of-Experts via Automated Pipeline Search

DGX agent

arXiv:2606.23739v1 Announce Type: cross Abstract: We present an automated large-scale search pipeline for heterogeneous 4-Expert Mixture-of-Experts (MoE4) architectures within the LEMUR neural network

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

T2D-Bench: Evidence-Gated Evaluation of LLM Outputs for Type 2 Diabetes Using a Multi-Layer Clinical-Lifestyle Knowledge Graph

DGX agent

arXiv:2606.24145v1 Announce Type: new Abstract: Large language models (LLMs) can produce clinically fluent recommendations for type 2 diabetes while failing to satisfy guideline constraints or explici

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Ten Digits on a Train: AI-Assisted Verification of Two Eigenvalue Problems

DGX agent

arXiv:2606.23821v1 Announce Type: cross Abstract: Accurate numerical eigenvalues are often difficult to certify, especially in singular or non-normal settings. This article reports a human--AI collabo

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

The African Language Tax: Quantifying the Cost, Latency, and Context Penalty of Tokenizing African Languages in Frontier LLMs

DGX agent

arXiv:2606.24460v1 Announce Type: cross Abstract: Commercial large language models bill, scale latency, and budget context per token. Yet tokenizers assign more subword tokens to the same meaning in s

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

The Degeneracy Distillery

DGX agent

arXiv:2606.23838v1 Announce Type: new Abstract: When two or more parameters or labels produce similar data, they are degenerate, or hard to distinguish. Degeneracies render both label prediction and i

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

TIGER: Taming Identity, Geometry, and Generative Priors for High-Quality Face Video Restoration

DGX agent

arXiv:2606.24336v1 Announce Type: new Abstract: Face Video Restoration (FVR) aims to recover high-fidelity facial videos from degraded input while preserving identity and semantic consistency across f

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

To Compare, or Not to Compare: On Methodological Practices in Evaluating Social Bias

DGX agent

arXiv:2606.24596v1 Announce Type: new Abstract: As Large Language Models are increasingly deployed in critical applications, robustly evaluating their social biases is paramount. However, the current

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian Sampling

DGX agent

arXiv:2606.24187v1 Announce Type: new Abstract: Long video understanding remains a daunting challenge for Multimodal Large Language Models (MLLMs) due to the excessive computation and memory footprint

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Towards Spec Learning: Inference-Time Alignment from Preference Pairs

DGX agent

arXiv:2606.24004v1 Announce Type: cross Abstract: Steering a large language model (LLM) toward a desired behavior typically relies on an iterative process of hand-crafting a prompt based on a careful

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Transformer-Based Language Models Across Domain Verticals: Architectures, Applications and Critical Assessment

DGX agent

arXiv:2606.24331v1 Announce Type: new Abstract: Transformer-based language models have become the default substrate for natural language processing and the pace of new releases has made it hard for pr

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Tri-Efficient Transfer Learning for Point Cloud Videos

DGX agent

arXiv:2606.24175v1 Announce Type: new Abstract: While point cloud foundation models have significantly advanced point cloud video understanding, existing parameter-efficient fine-tuning (PEFT) methods

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Trimming the Long-Tail of Visual World Modeling Evaluation

DGX agent

arXiv:2606.24256v1 Announce Type: new Abstract: Physical interactions follow a long-tailed distribution: a set of common and regular interactions dominates human experience and visual data, while a br

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

TrOCR for Medieval HTR: A Systematic Ablation Study with Cross-Dataset Validation

DGX agent

arXiv:2606.24302v1 Announce Type: new Abstract: Fine-tuning transformer-based handwritten text recognition (HTR) models on medieval manuscripts is challenging because these models are pre-trained on m

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Tuning without Peeking: Provable Generalization Bounds and Robust LLM Post-Training

DGX agent

arXiv:2507.01752v4 Announce Type: replace-cross Abstract: Gradient-based optimization is the workhorse of deep learning, offering efficient and scalable training via backpropagation. However, exposing

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving

DGX agent

arXiv:2606.24759v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong potential for autonomous driving scene understanding, yet existing methods still fac

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

UniRED: Unified RGB-D Video Frame Interpolation with Event Guidance

DGX agent

arXiv:2606.24282v1 Announce Type: new Abstract: High frame-rate RGB-D videos are crucial for a variety of downstream tasks, including motion analysis, dynamic scene understanding, and 3D reconstructio

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Variational Model Merging for Pareto Front Estimation in Multitask Finetuning

DGX agent

arXiv:2412.08147v2 Announce Type: replace-cross Abstract: Pareto fronts are useful to find good task-mixing strategies for multitask finetuning, but they are also costly to compute. To reduce costs, r

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

VeriPilot: An LLM-Powered Verilog Debugging Framework

DGX agent

arXiv:2606.23759v1 Announce Type: cross Abstract: Verilog debugging remains one of the most time-consuming stages in digital circuit design. Recent advances in Large Language Models (LLMs) have enable

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

video-SALMONN-R^3: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding

DGX agent

arXiv:2606.24477v1 Announce Type: cross Abstract: Video large language models (LLMs) are often constrained by computation and memory budgets, leading them to use reduced frame rates and spatial resolu

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

VisCritic: Visual State Comparison as Process Reward for GUI Agents

DGX agent

arXiv:2606.24525v1 Announce Type: new Abstract: GUI agents powered by vision-language models show strong potential for automating digital tasks, yet frequently fail in long-horizon scenarios due to th

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

What Does ODRL Mean? A Cross-Level Ontological Grounding of Permissions, Prohibitions, and Duties in UFO-L

DGX agent

arXiv:2606.24344v1 Announce Type: cross Abstract: ODRL policy evaluators produce verdicts, but say nothing about the normative positions a policy brings into existence, the authority structures those

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

When Helpfulness Overrides Causal Caution: Context-Dependent Suppression and Recovery in LLMs

DGX agent

arXiv:2606.24370v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into decision-support roles in business and policy contexts. While prior benchmark studies have

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

When Retrieval Metrics Mislead: Measuring Policy Signal in Long-Horizon Tool-Use Agents

DGX agent

arXiv:2606.23937v1 Announce Type: cross Abstract: Exact-match retrieval recall is often used as a proxy for whether a retriever supplies useful policy context to a downstream decision model. We test t

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

When Top-1 Fails: Calibrating LoRA Monitors for Masked Diffusion LMs

DGX agent

arXiv:2606.24119v1 Announce Type: cross Abstract: Discrete diffusion language model (DLM) fine-tuning inherits inexpensive diagnostics from denoising-time confidence monitors, but their PEFT-training

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

World Value Models for Robotic Manipulation

DGX agent

arXiv:2606.24742v1 Announce Type: new Abstract: Generalist value models play a pivotal role in scaling robotic policy learning from large-scale, mixed-quality data. Mathematically, accurate value esti

model-releasesarxiv-cs-ro
24 Jun 2026
Model Releases

You Don't Need to Run Every Eval

DGX agent

arXiv:2606.24020v1 Announce Type: new Abstract: A modern model release reports scores on 40+ benchmarks and the same evaluations were run many more times before it: to track training progress, compare

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

ZONOS2 Technical Report

DGX agent

arXiv:2606.24320v1 Announce Type: cross Abstract: We present ZONOS2 8B, our latest TTS model, which achieves state-of-the-art naturalness, prosody, and voice cloning fidelity. We improve upon Zonos-v0

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

4DVLT: Dynamic Scene Understanding with Worldline-Centered Vision-Language Tracking

DGX agent

arXiv:2606.22631v1 Announce Type: new Abstract: 4D dynamic scene understanding requires grounding language to a persistent worldline that binds identity, metric 3D motion, and synchronized multi-view

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

A-Evolve-Training: Autonomous Post-Training of a 30B Model

DGX agent

arXiv:2606.20657v1 Announce Type: cross Abstract: Post-training a frontier model is normally weeks of human work: proposing data and recipe changes, launching runs, reading evals, deciding what to kee

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

A Hybrid, Multi-Layered Pipeline for Phishing and Threat Classification: Independently Validated URL and NLP Engines with a Calibrated Multi-Channel Fusion Stage

DGX agent

arXiv:2606.21690v1 Announce Type: cross Abstract: Phishing is a multi-modal threat. We present a hybrid pipeline that scores each modality with its own engine and fuses the results. Three engines are

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

A Latent Representation Learning Framework for Hyperspectral Image Emulation in Remote Sensing

DGX agent

arXiv:2603.21911v2 Announce Type: replace Abstract: Synthetic hyperspectral image (HSI) generation is essential for large-scale simulation, algorithm development, and mission design, yet traditional r

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

A Linear Fractional Transformation Model and Calibration Method for Light Field Camera

DGX agent

arXiv:2511.03962v2 Announce Type: replace Abstract: Accurate intrinsic calibration is a crucial yet challenging prerequisite for 3D reconstruction using light field cameras. Existing calibration model

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

A Skin-Tone-Aware Dual-Representation Remote Photoplethysmography Framework for Contactless Respiratory Rate Estimation

DGX agent

arXiv:2606.21511v1 Announce Type: cross Abstract: Respiratory rate is a vital indicator of pulmonary and cardiovascular health, yet conventional methods for estimating respiratory rate are often intru

model-releasesarxiv-cs-cv
23 Jun 2026
← Previous
1…124125126127128…361
Next →