AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks

DGX agent

arXiv:2605.18583v1 Announce Type: cross Abstract: Coding agents now run autonomously with shell, file, and network privileges. When a user issues a benign request, the agent sometimes does more than a

model-releasesarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PaliBench: A Multi-Reference Blueprint for Classical Language Translation Benchmarks

DGX agent

arXiv:2605.16881v1 Announce Type: new Abstract: Digital humanities projects increasingly rely on machine translation and large language models to widen access to classical, religious, and otherwise un

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and Realistic Task Environments

DGX agent

arXiv:2603.23231v2 Announce Type: replace Abstract: Empowering large language models with long-term memory is crucial for building agents that adapt to users' evolving needs. Existing evaluations of t

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media

DGX agent

arXiv:2605.17187v1 Announce Type: cross Abstract: Social media are shifting towards pluralism -- community-governed platforms where groups define their own norms. What violates rules in one community

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Protection Is (Nearly) All You Need: Structural Protection Dominates Scoring in Globally Capped KV Eviction

DGX agent

arXiv:2605.18053v1 Announce Type: cross Abstract: We study KV cache eviction under a shared globally capped decode-time harness. Seven policies (LRU, H2O, SnapKV, StreamingLLM, Ada-KV, QUEST, Random)

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

ProtoSiTex: Learning Semi-Interpretable Prototypes for Multi-label Text Classification

DGX agent

arXiv:2510.12534v4 Announce Type: replace Abstract: The rapid growth of user-generated text across digital platforms has intensified the need for interpretable models capable of fine-grained text clas

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

SAM 2++: Tracking Anything at Any Granularity

DGX agent

arXiv:2510.18822v4 Announce Type: replace Abstract: Due to the varying granularity of target states across different tasks, most existing trackers are tailored to a single task, which specificity limi

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

SpecSem-Net: Integrating Spectral and Semantic Features for Robust AI-generated Video Detection

DGX agent

arXiv:2605.17311v1 Announce Type: new Abstract: The remarkable visual fidelity of recent commercial video generative models, such as Sora and Veo, renders robust AI-generated video detection increasin

model-releasesarxiv-cs-cv
19 May 2026
Hardware

Stable Audio 3

DGX agent

arXiv:2605.17991v1 Announce Type: cross Abstract: Stable Audio 3 is a family of fast latent diffusion models (small, medium, large) for variable-length audio generation and editing. Since our models c

hardwarearxiv-cs-ai
19 May 2026
Tutorials

StAD: Stein Amortized Divergence for Fast Likelihoods with Diffusion and Flow

DGX agent

arXiv:2605.16486v1 Announce Type: cross Abstract: Diffusion and flow-based models are ubiquitously used for generative modelling and density estimation. They admit a deterministic probability flow ord

tutorialsarxiv-cs-lg
19 May 2026
Model Releases

State-of-the-Art Claims Require State-of-the-Art Evidence

DGX agent

arXiv:2605.17273v1 Announce Type: cross Abstract: State-of-the-Art (SOTA) claims pervade Artificial Intelligence (AI) and Machine Learning (ML) research. These claims rest on benchmark evaluations, wh

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Supervising the search process produces reliable and generalizable information-seeking agents

DGX agent

arXiv:2502.13957v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are transforming web search by shifting from document ranking to synthesizing answers, and are increasingly deplo

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning

DGX agent

arXiv:2605.18109v1 Announce Type: new Abstract: In real home deployments, household agents must often operate from a complete household scene and a situated household request, rather than from a clean

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ToolMATH: A Diagnostic Benchmark for Long-Horizon Tool Use under Systematic Tool-Catalog Constraints

DGX agent

arXiv:2602.21265v2 Announce Type: replace Abstract: We introduce ToolMATH, a math-grounded diagnostic benchmark for evaluating long-horizon tool use under controllable tool-catalog conditions. ToolMAT

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey

DGX agent

arXiv:2409.10102v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has quickly grown into a pivotal paradigm in the development of Large Language Models (LLMs). Although ex

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TusoAI: Agentic Optimization for Scientific Methods

DGX agent

arXiv:2509.23986v2 Announce Type: replace Abstract: Scientific discovery is often slowed by the manual development of computational tools needed to analyze complex experimental data. Building such too

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation

DGX agent

arXiv:2605.18740v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) still struggle with fine-grained visual understanding, where answers often depend on small but decisive evide

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Visualizing the Invisible: Generative Visual Grounding Empowers Universal EEG Understanding in MLLMs

DGX agent

arXiv:2605.18172v1 Announce Type: new Abstract: Leveraging the universal representations of pre-trained LLMs and MLLMs offers a promising path toward brain foundation models. However, visually-evoked

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering

DGX agent

arXiv:2605.18313v1 Announce Type: cross Abstract: Small vision-language models (2-8B) are well-suited for clin- ical deployment due to privacy constraints, limited connectivity, and low-latency requir

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

When Personalization Legitimizes Risks: Uncovering Safety Vulnerabilities in Personalized Dialogue Agents

DGX agent

arXiv:2601.17887v2 Announce Type: replace Abstract: Long-term memory enables large language model (LLM) agents to support personalized and sustained interactions. However, most work on personalized ag

model-releasesarxiv-cs-ai
19 May 2026
Safety

Why Do Safety Guardrails Degrade Across Languages?

DGX agent

arXiv:2605.17173v1 Announce Type: cross Abstract: Large language models exhibit safety degradation in non-English languages. Standard evaluation relies on Jailbreak Success Rate (JSR), which confounds

safetyarxiv-cs-ai
19 May 2026
Model Releases

Beyond Bounded Variance: Variance-Reduced Normalized Methods for Nonconvex Optimization under Blum-Gladyshev Noise

DGX agent

arXiv:2605.15314v1 Announce Type: new Abstract: We study nonconvex stochastic optimization under the Blum-Gladyshev (mathsf{BG}-0) noise model, where the stochastic gradient variance grows quadratical

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

BiomedAP: A Vision-Informed Dual-Anchor Framework with Gated Cross-Modal Fusion for Robust Medical Vision-Language Adaptation

DGX agent

arXiv:2605.15736v1 Announce Type: cross Abstract: Biomedical Vision--Language Models (VLMs) have shown remarkable promise in few-shot medical diagnosis but face a critical bottleneck: extit{fragility

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Can We Trust AI-Inferred User States. A Psychometric Framework for Validating the Reliability of Users States Classification by LLMs in Operational Environments

DGX agent

arXiv:2605.15734v1 Announce Type: new Abstract: The use of large language models to assess user states in conversational and adaptive systems is based on the assumption that the metrics used for such

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

COPRA: Conditional Parameter Adaptation with Reinforcement Learning for Video Anomaly Detection

DGX agent

arXiv:2605.15325v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong performance in video anomaly detection (VAD) while providing interpretable predictions. However, existin

model-releasesarxiv-cs-cv
18 May 2026
Safety

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices

DGX agent

arXiv:2605.15684v1 Announce Type: new Abstract: The Diffusion Transformer (DiT) architecture is the state-of-the-art paradigm for high-fidelity image generation, underpinning models like Stable Diffus

safetyarxiv-cs-cv
18 May 2026
Model Releases

Fully Open Meditron: An Auditable Pipeline for Clinical LLMs

DGX agent

arXiv:2605.16215v1 Announce Type: new Abstract: Clinical decision support systems (CDSS) require scrutable, auditable pipelines that enable rigorous, reproducible validation. Yet current LLM-based CDS

model-releasesarxiv-cs-ai
18 May 2026
Research

GenAI-Driven Approach to RISC-V Supply Chain Exploration

DGX agent

arXiv:2605.15223v1 Announce Type: cross Abstract: This paper presents an LLM-empowered workflow for RISC-V supply chain analysis, integrating Vision-Language Models (VLMs) and Model-Driven Engineering

researcharxiv-cs-ai
18 May 2026
Model Releases

Harnessing Unimodality in Semiparametric Contextual Pricing via Oracle Price Map Learning

DGX agent

arXiv:2605.15411v1 Announce Type: cross Abstract: We study contextual dynamic pricing in a semiparametric scalar-index valuation model where the latent value is v_t=mu_ast(mathsf c_t)+xi_t, with an un

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Hidden in Memory: Sleeper Memory Poisoning in LLM Agents

DGX agent

arXiv:2605.15338v1 Announce Type: cross Abstract: Large language models are increasingly augmented with persistent memory, allowing assistants to store user-specific information across sessions for pe

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

IndicSafe: A Benchmark for Evaluating Multilingual LLM Safety in South Asia

DGX agent

arXiv:2603.17915v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are deployed in multilingual settings, their safety behavior in culturally diverse, low-resource languages rem

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

LoCO: Low-rank Compositional Rotation Fine-tuning

DGX agent

arXiv:2605.15916v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) has emerged as an critical technique for adapting large-scale foundation models across natural language process

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

MorphoHELM: A Comprehensive Benchmark for Evaluating Representations for Microscopy-Based Morphology Assays

DGX agent

arXiv:2605.15383v1 Announce Type: new Abstract: Microscopy images contain rich information about how cells respond to perturbations, making them essential to applications like drug screening. To quant

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Njord: A Probabilistic Graph Neural Network for Ensemble Ocean Forecasting

DGX agent

arXiv:2605.15470v1 Announce Type: new Abstract: Ocean dynamics are inherently chaotic, yet existing machine learning ocean models produce only deterministic forecasts. We introduce Njord, a probabilis

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control

DGX agent

arXiv:2605.15963v1 Announce Type: new Abstract: Large vision-language models have significantly advanced GUI agents, enabling executable interaction across web, mobile, and desktop interfaces. Yet the

model-releasesarxiv-cs-ai
18 May 2026
Applications

Process-Informed Forecasting of Complex Thermal Dynamics in Pharmaceutical Manufacturing

DGX agent

arXiv:2509.20349v3 Announce Type: replace Abstract: Accurate time-series forecasting for complex physical systems is the backbone of modern industrial monitoring and control, yet deep learning models

applicationsarxiv-cs-lg
18 May 2026
Model Releases

Quantum Feature Pyramid Gating for Seismic Image Segmentation

DGX agent

arXiv:2605.15370v1 Announce Type: cross Abstract: Accurate salt-body delineation is essential for seismic interpretation because salt structures distort wave propagation, complicate velocity-model bui

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

SaaS-Bench: Can Computer-Use Agents Leverage Real-World SaaS to Solve Professional Workflows?

DGX agent

arXiv:2605.15777v1 Announce Type: new Abstract: Computer-Using Agents (CUAs) are rapidly extending large language models (LLMs) beyond text-based reasoning toward action execution in more complex envi

model-releasesarxiv-cs-ai
18 May 2026
Local Ai

Sound Sparks Motion: Audio and Text Tuning for Video Editing

DGX agent

arXiv:2605.15307v1 Announce Type: cross Abstract: Motion-centric video editing remains difficult for large generative video models, which often respond well to appearance changes but struggle to produ

local-aiarxiv-cs-cv
18 May 2026
Model Releases

Structure-BiEval: A Self-Supervised, Dual-Track Framework for Decoupling Structure and Content in LLM Evaluation for Web Information Systems

DGX agent

arXiv:2601.19923v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into the core of Web-based autonomous agents and complex Web Information Systems, their ability to fait

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SynthRender and IRIS: Open-Source Framework and Dataset for Bidirectional Sim-Real Transfer in Industrial Object Perception

DGX agent

arXiv:2602.21141v2 Announce Type: replace Abstract: Object perception is fundamental for tasks such as robotic material handling and quality inspection. However, modern supervised deep-learning models

model-releasesarxiv-cs-cv
18 May 2026
Agents

Talking Trees: Reasoning-Assisted Induction of Decision Trees for Tabular Data

DGX agent

arXiv:2509.21465v3 Announce Type: replace Abstract: Tabular foundation models are becoming increasingly popular for low-resource tabular problems. These models make up for small training datasets by p

agentsarxiv-cs-lg
18 May 2026
Model Releases

TokenButler: Token Importance is Predictable

DGX agent

arXiv:2503.07518v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) rely on the Key-Value (KV) Cache to store token history, enabling efficient decoding of tokens. As the KV-Cache g

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

ACE-LoRA: Adaptive Orthogonal Decoupling for Continual Image Editing

DGX agent

arXiv:2605.14948v1 Announce Type: new Abstract: State-of-the-art diffusion models often rely on parameter-efficient fine-tuning to perform specialized image editing tasks. However, real-world applicat

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

ArGEnT: Arbitrary Geometry-encoded Transformer for Operator Learning

DGX agent

arXiv:2602.11626v2 Announce Type: replace-cross Abstract: Learning solution operators for systems with complex, varying geometries and parametric physical settings is a central challenge in scientific

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

AttnGen: Attention-Guided Saliency Learning for Interpretable Genomic Sequence Classification

DGX agent

arXiv:2605.14073v1 Announce Type: cross Abstract: Deep neural networks have achieved strong performance in genomic sequence classification; however, relating their predictions to biologically meaningf

model-releasesarxiv-cs-ai
15 May 2026
Research

CC-Pan: Channel-wise Compression based Diffusion for Efficient Pan-Sharpening

DGX agent

arXiv:2602.04473v2 Announce Type: replace Abstract: Recently, diffusion models have brought novel insights to pan-sharpening and notably boosted fusion precision. However, most existing models perform

researcharxiv-cs-cv
15 May 2026
Model Releases

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves

DGX agent

arXiv:2605.14068v1 Announce Type: new Abstract: We introduce CurveBench, a benchmark for hierarchical topological reasoning from visual input. CurveBench consists of extbf{756 images} of pairwise non-

model-releasesarxiv-cs-cv
15 May 2026
← Previous
1…368369370371372…1074
Next →