AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
15 May 2026

Phylogenetic Tree Inference with Tropical Axial Attention

SafetyDGX agent

arXiv:2605.13894v1 Announce Type: cross Abstract: In this work, we introduce a Tropical Axial Attention neural reasoning architecture that replaces vanilla softmax dot-product attention with max-plus

Prompt: Make a picture of a famous event in the history of humanity before smartphones existed, but in this picture people have smartphones and are all heads down obsessed with them

IndustryDGX agent

This Reddit post from r/ChatGPT shares a creative AI image generation prompt that asks an AI model to depict a famous historical event with a modern twist—showing people in that historical moment usin

RAG - Recherche Documentaire

Local AiDGX agent

Retrieval-Augmented Generation (RAG) is a technique that enhances large language models (LLMs) by enabling them to access and utilize external knowledge sources during response generation. This likely

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Realiz3D: 3D Generation Made Photorealistic via Domain-Aware Learning

TutorialsDGX agent

arXiv:2605.13852v1 Announce Type: cross Abstract: We often aim to generate images that are both photorealistic and 3D-consistent, adhering to precise geometry, material, and viewpoint controls. Typica

Reliability-Gated Source Anchoring for Continual Test-Time Adaptation

ResearchDGX agent

arXiv:2605.14063v1 Announce Type: new Abstract: Continual test-time adaptation (CTTA) updates a pretrained model online on an unlabeled, non-stationary stream while anchoring it to a frozen source che

SAGE3D: Soft-guided attention and graph excitation for 3D point cloud corner detection

ResearchDGX agent

arXiv:2605.15088v1 Announce Type: new Abstract: We present SAGE3D, a hybrid Transformer-based model for corner detection in airborne LiDAR point clouds. We propose a multi-stage solution built on a hi

ScaLoRA: Optimally Scaled Low-Rank Adaptation for Efficient High-Rank Fine-Tuning

ResearchDGX agent

arXiv:2510.23818v2 Announce Type: replace Abstract: As large language models (LLMs) continue to scale in size, the computational overhead has become a major bottleneck for task-specific fine-tuning. W

SCRWKV: Ultra-Compact Structure-Calibrated Vision-RWKV for Topological Crack Segmentation

ApplicationsDGX agent

arXiv:2605.14926v1 Announce Type: new Abstract: Achieving pixel-level accurate segmentation of structural cracks across diverse scenarios remains a formidable challenge. Existing methods face signific

Seed3D 2.0: Advancing High-Fidelity Simulation-Ready 3D Content Generation

Local AiDGX agent

arXiv:2605.13862v1 Announce Type: cross Abstract: We present Seed3D 2.0, an advanced 3D content generation system built on Seed3D 1.0, with substantial improvements across generation fidelity, simulat

Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility

Local AiDGX agent

arXiv:2605.14037v1 Announce Type: cross Abstract: Under modern test-time compute and agentic paradigms, language models process ever-longer sequences. Efficient text generation with transformer archit

SliceGraph: Mapping Process Isomers in Multi-Run Chain-of-Thought Reasoning

ResearchDGX agent

arXiv:2605.14619v1 Announce Type: new Abstract: Multi-run chain-of-thought reasoning is usually collapsed to final-answer aggregates, which discard howsampled trajectories share, split, and rejoin thr

Sort providers by cost, latency, or throughput on AI Gateway

ToolsDGX agent

Vercel's AI Gateway now includes functionality to sort AI providers based on performance metrics including cost, latency, and throughput. This feature enables developers to optimize their AI model sel

SR-Platform: An Agentic Pipeline for Natural Language-Driven Robot Simulation Environment Synthesis

AgentsDGX agent

arXiv:2605.14700v1 Announce Type: new Abstract: Generating robot simulation environments remains a major bottleneck in simulation-based robot learning. Constructing a training-ready MuJoCo scene typic

TeDiO: Temporal Diagonal Optimization for Training-Free Coherent Video Diffusion

ResearchDGX agent

arXiv:2605.14136v1 Announce Type: new Abstract: Recent text-to-video diffusion transformers generate visually compelling frames, yet still struggle with temporal coherence, often producing flickering,

Towards Continuous Sign Language Conversation from Isolated Signs

SafetyDGX agent

arXiv:2605.14705v1 Announce Type: new Abstract: Sign language is the primary language for many Deaf and Hard-of-Hearing (DHH) signers, yet most conversational AI systems still mediate interaction thro

TRIM: Token-wise Attention-Derived Saliency for Data-Efficient Instruction Tuning

ResearchDGX agent

arXiv:2510.07118v3 Announce Type: replace Abstract: Instruction tuning is essential for aligning large language models (LLMs) to downstream tasks and commonly relies on large, diverse corpora. However

Venus-DeFakerOne: Unified Fake Image Detection & Localization

Local AiDGX agent

arXiv:2605.14091v1 Announce Type: new Abstract: In recent years, the rapid evolution of generative AI has fundamentally reshaped the paradigm of image forgery, breaking the traditional boundaries betw

Video-Zero: Self-Evolution Video Understanding

ResearchDGX agent

arXiv:2605.14733v1 Announce Type: new Abstract: Self-evolution offers a promising path for improving reasoning models without relying on intensive human annotation. However, extending this paradigm to

Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining

AgentsDGX agent

arXiv:2605.14747v1 Announce Type: cross Abstract: Recent advances in multimodal large language models have driven growing interest in graphical user interface (GUI) agents, yet their generalization re

Vision-Core Guided Contrastive Learning for Balanced Multi-modal Prognosis Prediction of Stroke

SafetyDGX agent

arXiv:2605.14710v1 Announce Type: cross Abstract: Deep learning and multi-modal fusion have demonstrated transformative potential in medical diagnosis by integrating diverse data sources. However, acc

Volkswagen shows its first electric GTI; there's no chance of US sales

IndustryDGX agent

Volkswagen unveiled the ID. Polo GTI, its first electric GTI model, on May 15, 2026. The vehicle delivers 222 hp from a front-mounted electric motor and offers up to 282 miles of WLTP range. The ID. P

WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections

AgentsDGX agent

arXiv:2605.15030v1 Announce Type: cross Abstract: Web agents can autonomously complete online tasks by interacting with websites, but their exposure to open web environments makes them vulnerable to p

Widening the Gap: Exploiting LLM Quantization via Outlier Injection

Local AiDGX agent

arXiv:2605.15152v1 Announce Type: cross Abstract: LLM quantization has become essential for memory-efficient deployment. Recent work has shown that quantization schemes can pose critical security risk

You can now use your @grok subscription inside @NousResearch Hermes Agent. http://x.ai/news/grok-hermes

AgentsDGX agent

Nous Research has integrated Grok, xAI's large language model, into the Hermes Agent framework, allowing users with active @grok subscriptions to leverage Grok's capabilities within the Hermes Agent e

Your CLIP has 164 dimensions of noise: Exploring the embeddings covariance eigenspectrum of contrastively pretrained vision-language transformers

ResearchDGX agent

arXiv:2605.14893v1 Announce Type: cross Abstract: Contrastively pre-trained Vision-Language Models (VLMs) serve as powerful feature extractors. Yet, their shared latent spaces are prone to structural

14 May 2026

3D-UIR: 3D Gaussian for Underwater 3D Scene Reconstruction via Physics Based Appearance-Medium Decoupling

ResearchDGX agent

arXiv:2505.21238v3 Announce Type: replace Abstract: Novel view synthesis for underwater scene reconstruction presents unique challenges due to complex light-media interactions. Optical scattering and

A_3B_2: Adaptive Asymmetric Adapter for Alleviating Branch Bias in Vision-Language Image Classification with Few-Shot Learning

SafetyDGX agent

arXiv:2605.13161v1 Announce Type: new Abstract: Efficient transfer learning methods for large-scale vision-language models (e.g., CLIP) enable strong few-shot transfer, yet existing adaptation methods

Adaptive Conformal Prediction for Reliable and Explainable Medical Image Classification

SafetyDGX agent

arXiv:2605.12917v1 Announce Type: new Abstract: Deep learning models for medical imaging often exhibit overconfidence, creating safety risks in ambiguous diagnostic scenarios. While Conformal Predicti

AgentLens: Revealing The Lucky Pass Problem in SWE-Agent Evaluation

AgentsDGX agent

arXiv:2605.12925v1 Announce Type: cross Abstract: Evaluation of software engineering (SWE) agents is dominated by a binary signal: whether the final patch passes the tests. This outcome-only view trea

Anatomy-Slot: Unsupervised Anatomical Factorization for Homologous Bilateral Reasoning in Retinal Diagnosis

ResearchDGX agent

arXiv:2605.12929v1 Announce Type: cross Abstract: Retinal diagnosis is inherently bilateral: clinicians compare homologous structures across eyes (e.g., optic disc asymmetry), yet most deep models ope

Are Compact Rationales Free? Measuring Tile Selection Headroom in Frozen WSI-MIL

ResearchDGX agent

arXiv:2605.12575v1 Announce Type: cross Abstract: Whole-slide image (WSI) multiple instance learning (MIL) classifiers can achieve strong slide-level AUC while leaving the full-bag prediction opaque.

Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems

SafetyDGX agent

arXiv:2601.15161v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used for clinical decision support, where hallucinations and unsafe suggestions may pose direct

b9144

Local AiDGX agent

llama.cpp is an open source software library that performs inference on various large language models such as Llama. Build b9144 is an intermediate release version of the llama.cpp project from the gg

Byzantine-Robust Distributed Sparse Learning Revisited

ResearchDGX agent

arXiv:2605.13283v1 Announce Type: new Abstract: We revisit Byzantine robust distributed estimation for high-dimensional sparse linear models. By combining local ell_1-regularized robust estimation wit

CANTANTE: Optimizing Agentic Systems via Contrastive Credit Attribution

Local AiDGX agent

arXiv:2605.13295v1 Announce Type: cross Abstract: LLM-based multi-agent systems have demonstrated strong performance across complex real-world tasks, such as software engineering, predictive modeling,

Causal Learning with the Invariance Principle

ResearchDGX agent

arXiv:2605.13589v1 Announce Type: cross Abstract: Causal discovery, the problem of inferring the direction of causality, is generally ill-posed. We use the language of structural causal models (SCM) t

Chem-GMNet: A Sphere-Native Geometric Transformer for Molecular Property Prediction

ResearchDGX agent

arXiv:2605.13262v1 Announce Type: new Abstract: Modern SMILES-based chemical language models obtain strong MoleculeNet performance by treating SMILES as generic text and compensating with multi-millio

DeePen: Penetration Testing for Audio Deepfake Detection

ApplicationsDGX agent

arXiv:2502.20427v3 Announce Type: replace-cross Abstract: Deepfakes - manipulated or forged audio and video media - pose significant security risks to individuals, organizations, and society at large.

Diffusion-Inspired Reconfiguration of Transformers for Uncertainty Calibration

ResearchDGX agent

arXiv:2602.08920v2 Announce Type: replace Abstract: Uncertainty calibration in pre-trained transformers is critical for their reliable deployment in risk-sensitive applications. Yet, most existing pre

DisAgg: Distributed Aggregators for Efficient Secure Aggregation in Federated Learning

ResearchDGX agent

arXiv:2605.13708v1 Announce Type: cross Abstract: Federated learning enables collaborative model training across distributed clients, yet vanilla FL exposes client updates to the central server. Secur

Discrete Stochastic Localization for Non-autoregressive Generation

ResearchDGX agent

arXiv:2605.12836v1 Announce Type: new Abstract: Continuous diffusion is a natural framework for non-autoregressive generation but has generally lagged behind masked discrete diffusion models (MDMs) on

Diversity of Extensions in Abstract Argumentation

ResearchDGX agent

arXiv:2605.13332v1 Announce Type: new Abstract: Argumentation is an important topic of AI for modelling and reasoning about arguments. In abstract argumentation, we consider directed graphs, so-called

Do Activation Verbalization Methods Convey Privileged Information?

ResearchDGX agent

arXiv:2509.13316v4 Announce Type: replace-cross Abstract: Recent interpretability methods have proposed to translate LLM internal representations into natural language descriptions using a second verb

Filter-then-Weight: Online Data Selection and Reweighting for LLM Fine-Tuning

ResearchDGX agent

arXiv:2604.00001v2 Announce Type: replace-cross Abstract: Gradient-based data selection offers a principled framework for estimating sample utility in large language model (LLM) fine-tuning, but exist

Flow Matching for Offline Reinforcement Learning with Discrete Actions

SafetyDGX agent

arXiv:2602.06138v2 Announce Type: replace Abstract: Generative policies based on diffusion models and flow matching have shown strong promise for offline reinforcement learning (RL), but their applica

Flow Matching with Uncertainty Quantification and Guidance

ResearchDGX agent

arXiv:2602.10326v2 Announce Type: replace Abstract: Despite the remarkable success of sampling-based generative models such as flow matching, they can still produce samples of inconsistent or degraded

FRAME: Forensic Routing and Adaptive Multi-path Evidence Fusion for Image Manipulation Detection

ResearchDGX agent

arXiv:2605.12826v1 Announce Type: cross Abstract: The proliferation of sophisticated image editing tools and generative artificial intelligence models has made verifying the authenticity of digital im

GRACE: Gradient-aligned Reasoning Data Curation for Efficient Post-training

SafetyDGX agent

arXiv:2605.13130v1 Announce Type: new Abstract: Existing reasoning data curation pipelines score whole samples, treating every intermediate step as equally valuable. In reality, steps within a trace c

Graphon reels in $8.3M for its persistent relational memory platform

IndustryDGX agent

Graphon Inc., a startup with technology that makes artificial intelligence models better at processing large datasets, launched today with 8.3 million in funding. Novera Ventures led the seed round. I

HIR-ALIGN: Enhancing Hyperspectral Image Restoration via Diffusion-Based Data Generation

SafetyDGX agent

arXiv:2605.13581v1 Announce Type: new Abstract: Hyperspectral image (HSI) restoration is crucial for reliable analysis, as real HSIs suffer from degradations like noise, blur, and resolution loss. How

If you're calling a third-party API, your competitor can make the same call tomorrow. @lqiao is the keynote speaker at @pycon tomorrow, brea…

TutorialsDGX agent

If you're calling a third-party API, your competitor can make the same call tomorrow. @lqiao is the keynote speaker at @pycon tomorrow, breaking down how to build a durable moat using fine-tuned model

IGT-OMD: Implicit Gradient Transport for Decision-Focused Learning under Delayed Feedback

ResearchDGX agent

arXiv:2605.12693v1 Announce Type: new Abstract: Decision-focused learning trains predictive models end-to-end against downstream decision loss, but online settings suffer delayed feedback: outcomes ma

Language-Based Agent Control

AgentsDGX agent

arXiv:2605.12863v1 Announce Type: cross Abstract: This paper introduces language-based agent control (LBAC), a new programming model for agentic applications that brings techniques from programming la

Learning Transferable Latent User Preferences for Human-Aligned Decision Making

SafetyDGX agent

arXiv:2605.12682v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as reasoning modules in many applications. While they are efficient in certain tasks, LLMs often stru

Learning with Rare Success but Rich Feedback via Reflection-Enhanced Self-Distillation

SafetyDGX agent

arXiv:2605.12741v1 Announce Type: new Abstract: Enabling Large Language Models (LLMs) to continuously improve from environmental interactions is a central challenge in post-training. While on-policy s

LLMs as Implicit Imputers: Uncertainty Should Scale with Missing Information

ResearchDGX agent

arXiv:2605.13188v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in settings where the available context is incomplete or degraded. We argue that an LLM generat

Local Conformal Calibration of Dynamics Uncertainty from Semantic Images

Local AiDGX agent

arXiv:2605.13028v1 Announce Type: new Abstract: We introduce Observation-aware Conformal Uncertainty Local-Calibration (OCULAR), a conformal prediction-based algorithm that uses perception information

Looking for info on how to use reference images in A111 Forge Neo

TutorialsDGX agent

Forge Neo allows users to upload reference images to guide AI image generation, supporting JPG, PNG, and WebP formats. The platform supports various image generation and editing models including Flux,

MambaPanoptic: A Vision Mamba-based Structured State Space Framework for Panoptic Segmentation

ResearchDGX agent

arXiv:2605.12640v1 Announce Type: new Abstract: Panoptic segmentation requires the simultaneous recognition of countable thing instances and amorphous stuff regions, placing joint demands on long-rang

MARLIN: Multi-Agent Game-Theoretic Reinforcement Learning for Sustainable LLM Inference in Cloud Datacenters

AgentsDGX agent

arXiv:2605.13496v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become increasingly prevalent in cloud-based platforms, propelled by the introduction of AI-based consumer and enter

← Previous
1…774775776777778…1017
Next →