AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

ZEBRA: Zero-shot Budgeted Resource Allocation for LLM Orchestration

DGX agent

arXiv:2605.20485v1 Announce Type: new Abstract: As autonomous agents increasingly execute end-to-end tasks under fixed monetary budgets, the pressing open question shifts from whether the budget is re

model-releasesarxiv-cs-lg
21 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

A Bitter Lesson for Data Filtering

DGX agent

arXiv:2605.19407v1 Announce Type: cross Abstract: We investigate data filtering for large model pretraining via new scaling studies that target the high compute, data-scarce regime. In spite of an app

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Case for Agentic Tuning: From Documentation to Action in PostgreSQL

DGX agent

arXiv:2605.19988v1 Announce Type: cross Abstract: Documentation has long guided computer system tuning by distilling expert knowledge into per-parameter recommendations. Yet such guides capture only w

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Family of Divergence Measures for Evaluating the Reconstruction Quality of Explainable Ensemble Trees

DGX agent

arXiv:2605.19618v1 Announce Type: new Abstract: Validating interpretable surrogate models for ensemble learners requires measuring agreement between the ensemble's internal representation and its surr

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

A Hybrid Modeling Framework for Crop Prediction Tasks via Dynamic Parameter Calibration and Multi-Task Learning

DGX agent

arXiv:2603.15411v2 Announce Type: replace Abstract: Accurate prediction of crop states (e.g., phenology stages and cold hardiness) is essential for timely farm management decisions such as irrigation,

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Nonlinear Complexity Index for Wearable PPG Cardiovascular Stability: Multiscale Validation, Systematic Evaluation Correction, and Bayesian Parameter Optimization

DGX agent

arXiv:2605.18802v1 Announce Type: cross Abstract: Cardiovascular stability estimation from wearable photoplethysmography (PPG) requires a principled nonlinear framework, yet major gaps persist in heur

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Reproducibility Analysis of PO4ISR: Diagnosing and Mitigating Semantic Drift in LLM-Based Session Recommendation

DGX agent

arXiv:2605.18780v1 Announce Type: cross Abstract: Reasoning-based Large Language Models (LLMs) like PO4ISR have set new benchmarks in session-based recommendation. However, the reproducibility of thei

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Systematic Failure Analysis of Vision Foundation Models for Open Set Iris Presentation Attack Detection

DGX agent

arXiv:2605.19020v1 Announce Type: new Abstract: Vision foundation models have demonstrated strong transferability across diverse visual recognition tasks and are increasingly considered for biometric

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

A Two-Parameter Weibull Framework for Diagnosing Transformer Weight Distributions

DGX agent

arXiv:2605.18898v1 Announce Type: new Abstract: We apply the Weibull distribution -- a two-parameter family from extreme-value theory -- as a diagnostic framework for element-wise weight magnitude dis

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment

DGX agent

arXiv:2506.14148v2 Announce Type: replace-cross Abstract: This paper presents a novel non-invasive object classification approach using acoustic scattering, demonstrated through a case study on hair a

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Active Learning of Fractional-Order Viscoelastic Model Parameters for Realistic Haptic Rendering

DGX agent

arXiv:2512.00667v2 Announce Type: replace-cross Abstract: Effective medical simulators necessitate realistic haptic rendering of biological tissues that exhibit viscoelastic material properties, such

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

Adapted Center and Scale Prediction: More Stable and More Accurate

DGX agent

arXiv:2002.09053v3 Announce Type: replace Abstract: Pedestrian detection benefits from deep learning technology and gains rapid development in recent years. Most of detectors follow general object det

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Adaptive Power Iteration Method for Differentially Private PCA

DGX agent

arXiv:2602.11454v3 Announce Type: replace-cross Abstract: We study left(epsilon,eltaright)-differentially private algorithms for the problem of approximately computing the top singular vector of a mat

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Addressing prior dependence in hierarchical Bayesian modeling for PTA data analysis II: Noise and SGWB inference through parameter decorrelation

DGX agent

arXiv:2511.01959v2 Announce Type: replace-cross Abstract: Pulsar Timing Arrays (PTA) provide a powerful framework to measure low-frequency gravitational waves, but accuracy and robustness of the resul

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Adversarial Stress Testing of SPARK Humanoid Safety Filters

DGX agent

arXiv:2605.19009v1 Announce Type: new Abstract: Humanoid robots are difficult to deploy safely because they have high-dimensional bodies, many collision constraints, and must operate near people and o

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

Aero-World: Action-Conditioned Aerial Video Generation from Inertial Controls

DGX agent

arXiv:2605.19728v1 Announce Type: new Abstract: Foundation video models produce visually impressive results, but their use in embodied AI remains limited because they are primarily trained on natural

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Agent Meltdowns: The Road to Hell Is Paved with Helpful Agents

DGX agent

arXiv:2605.19149v1 Announce Type: new Abstract: Agents operating with computer and Web use inevitably encounter errors: inaccessible webpages, missing files, local and remote misconfigurations, etc. T

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

AgentNLQ: A General-Purpose Agent for Natural Language to SQL

DGX agent

arXiv:2605.19010v1 Announce Type: new Abstract: Natural language to SQL (NL2SQL) conversion is an important problem for researchers and enterprises due to the ubiquitous importance of relational datab

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

An Exterior Method for Nonnegative Matrix Factorization

DGX agent

arXiv:2605.19325v1 Announce Type: new Abstract: Nonnegative matrix factorization (NMF) seeks a low-rank approximation X approx UV^T with nonnegative factors and is commonly solved using interior metho

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

An LLM-Based System for Argument Mining

DGX agent

arXiv:2605.13793v2 Announce Type: replace Abstract: Arguments are a fundamental aspect of human reasoning, in which claims are supported, challenged, and weighed against one another. We present an end

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning

DGX agent

arXiv:2605.19852v1 Announce Type: new Abstract: Tool-augmented reasoning has emerged as a promising direction for enhancing the reasoning capabilities of multimodal large language models (MLLMs). Howe

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos

DGX agent

arXiv:2605.18984v1 Announce Type: new Abstract: Recent video generative models have greatly improved the realism of AI-generated videos, yet their outputs still exhibit artifacts such as temporal inco

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Auditing Reasoning-Trace Memorization Claims after Unlearning with Head-Conditioned Canaries

DGX agent

arXiv:2605.18891v1 Announce Type: cross Abstract: Evaluations of unlearning on reasoning models sometimes show a bypass pattern. The answer side looks unlearned, but the model's own thinking trace kee

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration

DGX agent

arXiv:2605.20025v1 Announce Type: new Abstract: Automating scientific discovery requires more than generating papers from ideas. Real research is iterative: hypotheses are challenged from multiple per

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Backdooring Masked Diffusion Language Models

DGX agent

arXiv:2605.19262v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are emerging as a compelling new paradigm for text generation, but their training-time security remains largely

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Base Models Look Human To AI Detectors

DGX agent

arXiv:2605.19516v1 Announce Type: cross Abstract: As AI-generated text enters the real-world at scale, institutions increasingly use commercial AI-text detectors, especially in education and academic-

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Benchmarking and Evolving Reason-Reflect-Rectify for Reflective Visual Generation

DGX agent

arXiv:2605.19639v1 Announce Type: new Abstract: Text-to-Image (T2I) models and Unified Multimodal Models (UMMs) have achieved remarkable progress in visual generation. However, their reliance on a sin

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Benchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and German

DGX agent

arXiv:2605.19069v1 Announce Type: cross Abstract: Code-switching -- the natural alternation between two languages within a single utterance -- represents one of the most challenging and under-studied

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Beyond Binary Success: A Diagnostic Meta-Evaluation Framework for Fine-Grained Manipulation

DGX agent

arXiv:2605.19986v1 Announce Type: cross Abstract: Fine-grained manipulation marks a regime where global scene context no longer suffices, and success hinges on the tight coupling of local attribute gr

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Beyond Imitation: Learning Safe End-to-End Autonomous Driving from Hard Negatives

DGX agent

arXiv:2605.19771v1 Announce Type: cross Abstract: Existing imitation learning methods for end-to-end autonomous driving predominantly learn from successful demonstrations by minimizing geometric devia

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Beyond Prediction Accuracy: Target-Space Recovery Profiles for Evaluating Model-Brain Alignment

DGX agent

arXiv:2605.20127v1 Announce Type: cross Abstract: Artificial vision models are often evaluated against the human visual cortex by measuring how accurately their internal representations predict brain

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Beyond Waypoints: Dual-Heatmap Grounding for Cross-Embodiment Semantic Navigation

DGX agent

arXiv:2605.19420v1 Announce Type: new Abstract: Grounding open-ended semantic instructions into physically executable local goals is a fundamental challenge in human-robot interaction. While existing

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

BLINKG: A Benchmark for LLM-Integrated Knowledge Graph Generation

DGX agent

arXiv:2605.19518v1 Announce Type: new Abstract: Generating Knowledge Graphs (KGs) remains one of the most time-consuming and labor-intensive tasks for knowledge engineers, as they need to identify sem

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

BuildArena: A Physics-Aligned Interactive Benchmark of LLMs for Engineering Construction

DGX agent

arXiv:2510.16559v5 Announce Type: replace Abstract: Engineering construction automation aims to transform natural language specifications into physically viable structures, requiring complex integrate

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Can Large Language Models Reliably Correct Errors in Low-Resource ASR? A Contamination-Aware Case Study on West Frisian

DGX agent

arXiv:2605.19711v1 Announce Type: new Abstract: Automatic speech recognition (ASR) has improved substantially in recent years, yet performance remains limited for low-resource languages. Large languag

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Can LLMs Emulate Human Belief Dynamics?

DGX agent

arXiv:2605.18781v1 Announce Type: cross Abstract: Can LLMs simulate how humans form and change beliefs in social networks? We put this to the test by replicating an established study on belief dynamic

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision

DGX agent

arXiv:2605.19538v1 Announce Type: cross Abstract: CAPTCHAs are widely deployed as human verification mechanisms and frequently block intelligent agents from completing end-to-end automation in real-wo

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Causal Evidence for Attention Head Imbalance in Modality Conflict Hallucination

DGX agent

arXiv:2605.19250v1 Announce Type: new Abstract: Modality-conflict hallucination occurs when multimodal large language models (MLLMs) prioritize erroneous textual premises over contradictory visual evi

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Chunking German Legal Code

DGX agent

arXiv:2605.19806v1 Announce Type: cross Abstract: This paper investigates chunking strategies for retrieval-augmented generation on German statutory law, using the German Civil Code as a structured be

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning

DGX agent

arXiv:2605.20176v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems have shown promise for clinical decision support, but existing works largely assume that evidence has a

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

ClusterRAG: Cluster-Based Collaborative Filtering for Personalized Retrieval-Augmented Generation

DGX agent

arXiv:2605.18769v1 Announce Type: cross Abstract: Personalized Retrieval-Augmented Generation (RAG) relies on accurately selecting user-relevant documents. In practice, existing RAG approaches often s

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

CogScale: Scalable Benchmark for Sequence Processing

DGX agent

arXiv:2605.19758v1 Announce Type: new Abstract: The ability to maintain and manipulate information over time is a fundamental aspect of living beings and Artificial Intelligence. While modern models h

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

COMPASS: Confined-space Manipulation Planning with Active Sensing Strategy

DGX agent

arXiv:2509.14787v2 Announce Type: replace Abstract: Manipulation in confined and cluttered environments remains a significant challenge due to partial observability and complex configuration spaces. E

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect

DGX agent

arXiv:2605.18808v1 Announce Type: cross Abstract: We characterize a compositional architecture of literary primitives in two instruction-tuned large language models (Llama 3.1 8B-Instruct and Gemma 2

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Conflict-Resilient Multi-Agent Reasoning via Signed Graph Modeling

DGX agent

arXiv:2605.19418v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) have demonstrated strong reasoning and decision-making capabilities that consistently surpass those of single LLM ag

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations

DGX agent

arXiv:2603.17305v2 Announce Type: replace Abstract: We propose CRAFT, a red-teaming alignment framework that leverages model reasoning capabilities and hidden representations to improve robustness aga

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering

DGX agent

arXiv:2605.19075v1 Announce Type: cross Abstract: Grounded multi-video question answering over real-world news events requires systems to surface query-relevant evidence across heterogeneous video arc

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Cross-View Attention Fusion Net: A Prior-Guided Dual-View Representation Learning for Cardiac Output Estimation from Short-Term PPG Signals

DGX agent

arXiv:2605.19666v1 Announce Type: cross Abstract: Accurate cardiac output (CO) estimation from photoplethysmography (PPG) is promising for unobtrusive hemodynamic monitoring, but remains difficult sin

model-releasesarxiv-cs-lg
20 May 2026
← Previous
1…216217218219220…361
Next →