AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

PetroBench: A Benchmark for Large Language Models in Petroleum Engineering

DGX agent

arXiv:2605.28032v1 Announce Type: new Abstract: Large Language Models are increasingly applied in the petroleum industry, highlighting the need for a domain-specific evaluation framework. This study d

model-releasesarxiv-cs-ai
28 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PINE: Pruning Boosted Tree Ensembles with Conformal In-Distribution Prediction Equivalence

DGX agent

arXiv:2605.28068v1 Announce Type: new Abstract: Tree ensembles are machine learning models with strong predictive performance and interpretability, and remain widely used for tabular data. Standard pr

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona (Rashi Shrivastava/Forbes)

DGX agent

Rashi Shrivastava / Forbes: Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona — Gray Swan works

model-releasestechmeme
28 May 2026
Model Releases

Plant, Persist, Trigger: Sleeper Attack on Large Language Model Agents

DGX agent

arXiv:2605.28201v1 Announce Type: new Abstract: Large Language Model (LLM) agents remain vulnerable to safety threats from the external environment, where attackers inject adversarial content into ext

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control

DGX agent

arXiv:2601.15015v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown promising results in active flow control (AFC), yet progress in the field remains difficult to assess as exist

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation

DGX agent

arXiv:2605.28237v1 Announce Type: cross Abstract: Real-world navigation is fundamentally driven by Points of Interest (POIs), yet reaching a precise POI remains a critical 'final-meters' challenge. Ex

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

PointQ-Bench: Benchmarking Diagnostic and Interpretable Point Cloud Quality Assessment

DGX agent

arXiv:2605.28241v1 Announce Type: new Abstract: Point cloud quality plays a critical role in 3D acquisition, reconstruction, rendering, and perception, yet existing point cloud quality assessment (PCQ

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

PortBench: A Correlation-Aware, Full-Pipeline Benchmark for LLM-Driven Portfolio Management

DGX agent

arXiv:2605.27887v1 Announce Type: new Abstract: LLMs have shown strong performance across diverse financial tasks, yet portfolio management (PM), a critical financial decision-making task, remains poo

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Pressure-Testing Deception Probes in LLMs: Scaling, Robustness, and the Geometry of Deceptive Representations

DGX agent

arXiv:2605.27958v1 Announce Type: cross Abstract: Linear probes trained on LLM activations are increasingly proposed as deception-detection metrics, yet report AUROC exceeding 0.96 on clean benchmarks

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

PrionNER: A Named Entity Recognition Dataset for Prion Disease Biomedical Literature

DGX agent

arXiv:2605.28375v1 Announce Type: new Abstract: Prion diseases are rare, rapidly progressive, and fatal neurodegenerative disorders that remain difficult to diagnose, particularly in their early stage

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Privately Estimating Monotone Statistics in Polynomial Time

DGX agent

arXiv:2605.27912v1 Announce Type: cross Abstract: We study efficient differentially private algorithms for estimating monotone statistics, i.e., statistics that are monotone under the addition of new

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Probabilistic Data-Driven Modelling of Astrophysical Transients: The Neural Process Family for Ultrafast and Class-Agnostic Light Curve Reconstruction with NightLANP

DGX agent

arXiv:2605.27527v1 Announce Type: cross Abstract: Astrophysical observations taken from Earth are subject to weather, environmental, and scientific constraints that lead to sparse, irregular light cur

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Probing for Knowledge Attribution in Large Language Models

DGX agent

arXiv:2602.22787v2 Announce Type: replace-cross Abstract: Large language model (LLM) hallucinations, meaning fluent but factually incorrect generations, fall into two types: faithfulness violations, w

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

ProgVLA: Progress-Aware Robot Manipulation Skill Learning

DGX agent

arXiv:2605.28231v1 Announce Type: cross Abstract: We present ProgVLA, a compact vision-language-action (VLA) model designed for reliable robot manipulation under tight compute and memory budgets. The

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Prominence-Stratified Failure Modes in Retrieval-Augmented Commercial Recommendation: A 37,000-Run Audit

DGX agent

arXiv:2605.27439v1 Announce Type: cross Abstract: AI assistants like ChatGPT and Claude are recommendation engines, not search engines: they answer commercial queries by directly nominating brands rat

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Prompt Codebooks: Discrete Compositional Optimization for Language Model Instruction Refinement

DGX agent

arXiv:2605.28360v1 Announce Type: new Abstract: Automatic prompt optimization (APO) has driven significant gains in LLM-based agentic workflows. However, existing methods treat each task's prompt as a

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

PromptEmbedder:: Efficient and Transferable Text Embedding via Dual-LLM Soft Prompting

DGX agent

arXiv:2605.28066v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable efficacy in text embedding, yet current adaptation methods like LoRA face significant bottle

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Prompting Is All You Need: Multi-view Prompting Large Language Models for Aspect-Based Sentiment Analysis

DGX agent

arXiv:2605.28058v1 Announce Type: new Abstract: Recent work explored the capabilities of Large Language Models (LLMs) in Aspect-Based Sentiment Analysis (ABSA) through few-shot prompting, requiring su

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

ProvMind: Provenance-grounded reasoning for materials synthesis

DGX agent

arXiv:2605.28487v1 Announce Type: new Abstract: Materials process optimization requires reasoning over routes, conditions, tools and causal dependencies, yet most computational formulations flatten sy

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

PrunePath: Towards Highly Structured Sparse Language Models

DGX agent

arXiv:2605.28283v1 Announce Type: cross Abstract: Feed-forward networks (FFNs) dominate the parameter count and computation of modern language models, yet existing pruning methods often struggle to co

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Pruning and Distilling Mixture-of-Experts into Dense Language Models

DGX agent

arXiv:2605.28207v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) is now the dominant architecture for frontier language models, yet it requires all expert parameters to be loaded in memory,

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

PubMedCausal: A Span-Level Annotated Corpus for Causal Relation Extraction in Biomedical Text

DGX agent

arXiv:2605.28363v1 Announce Type: new Abstract: Causal relation extraction (CRE) is central to biomedical text mining, but current resources often conflate causal relations with broader associations,

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation

DGX agent

arXiv:2605.28091v1 Announce Type: new Abstract: Text-to-Image generation has evolved from basic image synthesis into a frequently used core capability in professional creative workflows, where simple

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

📢Qwen3.7-Max just hit #3 on ITbench-AA — a fresh benchmark testing how well models handle real-world enterprise IT tasks, agentic-style. 🔧…

DGX agent

📢Qwen3.7-Max just hit #3 on ITbench-AA — a fresh benchmark testing how well models handle real-world enterprise IT tasks, agentic-style. 🔧Agentic era, go with Qwen.🏃🏃 Artificial Analysis and IBM Resea

model-releasesqwen--x
28 May 2026
Model Releases

RASR: Retrieval-Augmented Super Resolution for Practical Reference-based Image Restoration

DGX agent

arXiv:2508.09449v2 Announce Type: replace Abstract: Reference-based Super Resolution (RefSR) improves upon Single Image Super Resolution (SISR) by leveraging high-quality reference images to enhance t

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

REED: Post-Training Representation Editing for Cross-Domain Linguistic Steganalysis

DGX agent

arXiv:2605.28298v1 Announce Type: new Abstract: In real-world scenarios of linguistic steganalysis, tested texts usually come from unseen domains with different vocabularies, topics, writing styles, a

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Reevaluating Policy Gradient Methods for Imperfect-Information Games

DGX agent

arXiv:2502.08938v4 Announce Type: replace Abstract: In the past decade, motivated by the putative failure of naive self-play deep reinforcement learning (DRL) in adversarial imperfect-information game

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Reflective Dialogue between Teacher and Solver Agents for Video Question Answering

DGX agent

arXiv:2605.27885v1 Announce Type: new Abstract: Various approaches have been proposed to adapt Vision-Language Models (VLMs) to specialized domains for Video Question Answering, including fine-tuning

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

ReflexGrad: Within-Episode Failure Recovery in LLM Agents via Progress-Gated Dual-Process Routing

DGX agent

arXiv:2511.14584v3 Announce Type: replace-cross Abstract: We present ReflexGrad, a dual-process architecture for within-episode failure recovery in LLM agents without demonstrations. When agents commi

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Regression Language Models for Code

DGX agent

arXiv:2509.26476v2 Announce Type: replace-cross Abstract: We study code-to-metric regression: predicting numeric outcomes of code executions, a challenging task due to the open-ended nature of program

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Relational Semantic Reasoning on 3D Scene Graphs for Open World Interactive Object Search

DGX agent

arXiv:2603.05642v2 Announce Type: replace-cross Abstract: Open-world interactive object search in household environments requires understanding semantic relationships between objects and their surroun

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG

DGX agent

arXiv:2605.28044v1 Announce Type: new Abstract: Cited RAG evaluation often treats visible sources as a grounding signal, but a real, topically relevant citation can still under-warrant the attached wo

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

ReSAE: Residualized Sparse Autoencoders for Multi-Layer Transformer Interventions

DGX agent

arXiv:2605.27819v1 Announce Type: cross Abstract: Sparse autoencoders are usually trained one layer at a time, even though transformer residual stream activations are strongly coupled across depth. Th

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Resolution-free neural surrogates for geometric parameterization and mapping with spatially varying fields

DGX agent

arXiv:2605.28551v1 Announce Type: new Abstract: Many imaging problems require computing spatial transformations induced by spatially varying intensity, feature, or density fields. Canonical examples i

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Resource-Constrained Affect Modelling via Variance Regularisation Pruning

DGX agent

arXiv:2605.27479v1 Announce Type: cross Abstract: Affective computing systems are increasingly embedded in pervasive and interactive environments, such as adaptive games, assistive technologies, and r

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Revisiting 2D Foundation Models for Scalable 3D Medical Image Classification

DGX agent

arXiv:2512.12887v3 Announce Type: replace Abstract: 3D medical image classification is essential for modern clinical workflows. Medical foundation models (FMs) have emerged as a promising approach for

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Revisiting Metafeatures to Explain Model Differences on Tabular Data

DGX agent

arXiv:2605.28418v1 Announce Type: new Abstract: With the rise of tabular foundation models alongside traditional models still performing well on many tasks, choosing the right model for a tabular data

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

RGC: a radio AGN classifier based on deep learning. I. A semi-supervised multiclass model for VLA images

DGX agent

arXiv:2510.22190v2 Announce Type: replace-cross Abstract: Bent radio active galactic nuclei (RAGNs) -- wide-angle tails (WATs) and narrow-angle tails (NATs) -- trace dense environments in galaxy group

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

RMPL: Relation-aware Multi-task Progressive Learning with Stage-wise Training for Multimedia Event Extraction

DGX agent

arXiv:2602.13748v2 Announce Type: replace Abstract: Multimedia Event Extraction (MEE) aims to identify events and their arguments from documents that contain both text and images. It requires groundin

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Robust Moment-Based Estimation via Spectral Gradient Reweighting

DGX agent

arXiv:2605.27718v1 Announce Type: cross Abstract: Moment-based estimation is a theoretically attractive approach to parametric inference, especially when likelihood-based estimation is unavailable, mi

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

RW-TTT: Batched Serving for Request-Owned Test-Time Training State

DGX agent

arXiv:2605.28053v1 Announce Type: new Abstract: Test-time training (TTT) adapts an LLM during generation by reading and updating request-owned state, such as fast weights, low-rank deltas, or streamin

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Safe In-Context Reinforcement Learning

DGX agent

arXiv:2509.25582v3 Announce Type: replace Abstract: In-context reinforcement learning (ICRL) is an emerging RL paradigm where an agent, after pretraining, can adapt to out-of-distribution test tasks w

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

SAM-Enhanced Segmentation on Road Datasets: Balancing Critical Classes in Autonomous Driving

DGX agent

arXiv:2605.28136v1 Announce Type: new Abstract: Dense semantic segmentation is essential for autonomous driving, yet many multi-modal datasets lack pixel-level annotations. The Zenseact Open Dataset (

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2602.01990v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance through instruction tuning, but real-world deployment requires them to con

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

SeeGroup: Multi-Layer Depth Estimation of Transparent Surfaces via Self-Determined Grouping

DGX agent

arXiv:2605.28735v1 Announce Type: new Abstract: Transparent objects are common in daily life, and it is important to understand their multilayer depth, including the transparent surface and the object

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Self-Supervised Online Robot-Agnostic Traversability Estimation for Open-World Environments

DGX agent

arXiv:2605.28442v1 Announce Type: cross Abstract: Self-supervised online traversability estimation enables robots to continuously learn from unlabeled open-world experiences and adapt their navigation

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

SHIPPED. Mistral Vibe is now the AI agent for long-horizon productivity and coding, and the home for Work mode, Code mode, the CLI, and a br…

DGX agent

Mistral AI has released Mistral Vibe, an AI agent designed for long-horizon productivity and coding tasks, featuring Work mode, Code mode, a CLI, and additional capabilities. The product consolidates

model-releasesmistral-ai--x
28 May 2026
Model Releases

SIGMA: Bridging Structural and Distributional Gaps for Vision Foundation Model Adaptation

DGX agent

arXiv:2605.27893v1 Announce Type: new Abstract: Vision Foundation Models (VFMs) have demonstrated impressive representational capabilities. However, adapting them to downstream tasks via full fine-tun

model-releasesarxiv-cs-cv
28 May 2026
← Previous
1…256257258259260…475
Next →