AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,690 results
Safety

Mitigating Misalignment Contagion by Steering with Implicit Traits

DGX agent

arXiv:2605.02751v1 Announce Type: cross Abstract: Language models (LMs) are increasingly used in high-stakes, multi-agent settings, where following instructions and maintaining value alignment are cri

safetyarxiv-cs-cl
5 May 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Mitigating Multimodal LLMs Hallucinations via Relevance Propagation at Inference Time

DGX agent

arXiv:2605.01766v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have revolutionized the landscape of AI, demonstrating impressive capabilities in tackling complex vision and

researcharxiv-cs-cv
5 May 2026
Model Releases

MOSAIC: Multi-agent Orchestration for Task-Intelligent Scientific Coding

DGX agent

arXiv:2510.08804v3 Announce Type: replace Abstract: We present MOSAIC, a multi-agent Large Language Model (LLM) framework for solving challenging scientific coding tasks. Unlike general-purpose coding

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety

DGX agent

arXiv:2605.01687v1 Announce Type: new Abstract: We present MultiBreak, a scalable and diverse multi-turn jailbreak benchmark to evaluate large language model (LLM) safety. Multi-turn jailbreaks mimic

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Pandora's Regret: A Proper Scoring Rule for Evaluating Sequential Search

DGX agent

arXiv:2605.01936v1 Announce Type: new Abstract: In sequential search, alternatives are tested until the true class is found. Standard proper scoring rules like log loss are local, ignoring the ranking

model-releasesarxiv-cs-lg
5 May 2026
Research

ParaRNN: An Interpretable and Parallelizable Recurrent Neural Network for Time-Dependent Data

DGX agent

arXiv:2605.02692v1 Announce Type: cross Abstract: The proliferation of large-scale and structurally complex data has spurred the integration of machine learning methods into statistical modeling. Recu

researcharxiv-cs-lg
5 May 2026
Model Releases

Physics-Informed Neural Learning for State Reconstruction and Parameter Identification in Coupled Greenhouse Climate Dynamics

DGX agent

arXiv:2605.02524v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have recently emerged as a promising framework for integrating data-driven learning with physical knowledge. In

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Prescriptive Scaling Laws for Data Constrained Training

DGX agent

arXiv:2605.01640v1 Announce Type: cross Abstract: Training compute is increasingly outpacing the availability of high-quality data. This shifts the central challenge from optimal compute allocation to

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Probe-Geometry Alignment: Erasing the Cross-Sequence Memorization Signature Below Chance

DGX agent

arXiv:2605.01699v1 Announce Type: new Abstract: Recent attacks show that behavioural unlearning of large language models leaves internal traces recoverable by adversarial probes. We characterise where

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Projection-Free Transformers via Gaussian Kernel Attention

DGX agent

arXiv:2605.02144v1 Announce Type: new Abstract: Self-attention in Transformers is typically implemented as softmax(QK^op/sqrt{d})V, where Q=XW_Q, K=XW_K, and V=XW_V are learned linear projections of t

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces

DGX agent

arXiv:2605.02801v1 Announce Type: new Abstract: As large language model (LLM) agents evolve from isolated tool users into coordinated teams, reinforcement learning (RL) must optimize not only individu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Rethinking Multi-Label Node Classification: Do Tuned Classic GNNs Suffice?

DGX agent

arXiv:2605.01403v1 Announce Type: new Abstract: Multi-label node classification (MLNC) has recently been addressed by increasingly complex label-aware designs that explicitly model node-label interact

model-releasesarxiv-cs-lg
5 May 2026
Research

Rhamba: Region-Aware Hybrid Attention-Mamba Framework for Self-Supervised Learning in Resting-State fMRI

DGX agent

arXiv:2605.01240v1 Announce Type: new Abstract: Self-supervised pretraining is promising for large-scale neuroimaging, yet the impact of region-aware masking and hybrid sequence modeling remains under

researcharxiv-cs-lg
5 May 2026
Agents

S^3-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data

DGX agent

arXiv:2605.01248v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has enabled newer capabilities in models, such as agentic tool-use for search. However, these models struggle

agentsarxiv-cs-lg
5 May 2026
Model Releases

SAMamba3D: adapting Segment Anything for generalizable 3D segmentation of multiphase pore-scale images

DGX agent

arXiv:2605.00916v1 Announce Type: new Abstract: Reliable segmentation of multiphase pore-scale X-ray images of rocks is necessary to quantify fluid saturation, connectivity, and interfacial geometry.

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

SemEval-2026 Task 7: Everyday Knowledge Across Diverse Languages and Cultures

DGX agent

arXiv:2605.02601v1 Announce Type: new Abstract: We present our shared task on evaluating the adaptability of LLMs and NLP systems across multiple languages and cultures. The task data consist of an ex

model-releasesarxiv-cs-cl
5 May 2026
Safety

Skills as Verifiable Artifacts: A Trust Schema and a Biconditional Correctness Criterion for Human-in-the-Loop Agent Runtimes

DGX agent

arXiv:2605.00424v1 Announce Type: cross Abstract: Agent skills -- structured packages of instructions, scripts, and references that augment a large language model (LLM) without modifying the model its

safetyarxiv-cs-ai
5 May 2026
Model Releases

STEP: Warm-Started Visuomotor Policies with Spatiotemporal Consistency Prediction

DGX agent

arXiv:2602.08245v2 Announce Type: replace Abstract: Diffusion policies have recently emerged as a powerful paradigm for visuomotor control in robotic manipulation due to their ability to model the dis

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

SwiftChannel: Algorithm-Hardware Co-Design for Deep Learning-Based 5G Channel Estimation

DGX agent

arXiv:2605.01931v1 Announce Type: cross Abstract: Channel estimation is crucial in 5G communication networks for optimizing transmission parameters and ensuring reliable, high-speed communication. How

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Task-Driven Subspace Decomposition for Knowledge Sharing and Isolation in LoRA-based Continual Learning

DGX agent

arXiv:2603.00191v2 Announce Type: replace-cross Abstract: Continual Learning (CL) requires models to sequentially adapt to new tasks without forgetting old knowledge. Recently, Low-Rank Adaptation (Lo

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Teaching LLMs Brazilian Healthcare: Injecting Knowledge from Official Clinical Guidelines

DGX agent

arXiv:2605.01077v1 Announce Type: new Abstract: Brazil's Unified Health System (SUS) relies on official clinical guidelines that define diagnostic criteria, treatments, dosages, and monitoring procedu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

The Good, the Bad, and the Sampled: a No-Regret Approach to Safe Online Classification

DGX agent

arXiv:2510.01020v2 Announce Type: replace Abstract: We study sequential testing for a binary disease outcome when risk follows an unknown logistic model. At each round, the decision maker may either p

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Towards Lightest Low-Light Image Enhancement Architecture for Mobile Devices

DGX agent

arXiv:2507.04277v2 Announce Type: replace Abstract: Real-time low-light image enhancement on mobile and embedded devices requires models that balance visual quality and computational efficiency. Exist

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

VILAS: A VLA-Integrated Low-cost Architecture with Soft Grasping for Robotic Manipulation

DGX agent

arXiv:2605.02037v1 Announce Type: new Abstract: We present VILAS, a fully low-cost, modular robotic manipulation platform designed to support end-to-end vision-language-action (VLA) policy learning an

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Borrowed Geometry: Computational Reuse of Frozen Text-Pretrained Transformer Weights Across Modalities

DGX agent

arXiv:2605.00333v1 Announce Type: cross Abstract: Frozen Gemma 4 31B weights pretrained exclusively on text tokens, unmodified, transfer across modality boundaries through a thin trainable interface.

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection

DGX agent

arXiv:2605.00630v1 Announce Type: new Abstract: The proliferation of advanced AI video synthesis techniques poses an unprecedented challenge to digital video authenticity. Existing AI-generated video

model-releasesarxiv-cs-cv
4 May 2026
Research

Confidence Estimation in Automatic Short Answer Grading with LLMs

DGX agent

arXiv:2605.00200v1 Announce Type: new Abstract: Automatic Short Answer Grading (ASAG) with generative large language models (LLMs) has recently demonstrated strong performance without task-specific fi

researcharxiv-cs-cl
4 May 2026
Model Releases

Deepfakes: we need to re-think the concept of 'real' images

DGX agent

arXiv:2509.21864v2 Announce Type: replace Abstract: The wide availability and low usability barrier of modern image generation models has triggered the reasonable fear of criminal misconduct and negat

model-releasesarxiv-cs-cv
4 May 2026
Research

End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer

DGX agent

arXiv:2605.00503v1 Announce Type: new Abstract: Autoregressive image modeling relies on visual tokenizers to compress images into compact latent representations. We design an end-to-end training pipel

researcharxiv-cs-cv
4 May 2026
Model Releases

FedACT: Concurrent Federated Intelligence across Heterogeneous Data Sources

DGX agent

arXiv:2605.00011v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative intelligence across decentralized data source devices in a privacy-preserving way. While substantial resea

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios

DGX agent

arXiv:2605.00706v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied in financial scenarios. However, they may produce harmful outputs, including facilitating illegal

model-releasesarxiv-cs-cl
4 May 2026
Research

Flow matching for Sentinel-2 super-resolution: implementation, application, and implications

DGX agent

arXiv:2605.00367v1 Announce Type: new Abstract: Developing robust techniques for super-resolution of satellite imagery involves navigating commonly observed trade-offs between spectral fidelity and pe

researcharxiv-cs-cv
4 May 2026
Safety

FreeRet: MLLMs as Training-Free Retrievers

DGX agent

arXiv:2509.24621v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are emerging as versatile foundations for mixed-modality retrieval. Yet, they often require heavy post-hoc

safetyarxiv-cs-cv
4 May 2026
Model Releases

Hypergraph and Latent ODE Learning for Multimodal Root Cause Localization in Microservices

DGX agent

arXiv:2605.00351v1 Announce Type: new Abstract: Root cause localization in cloud native microservice systems requires modeling complex service dependencies, irregular temporal dynamics, and heterogene

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Learning from the Unseen: Generative Data Augmentation for Geometric-Semantic Accident Anticipation

DGX agent

arXiv:2605.00051v1 Announce Type: new Abstract: Anticipating traffic accidents is a critical yet unresolved problem for autonomous driving, hindered by the inherent complexity of modeling interactions

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

LLM-Oriented Information Retrieval: A Denoising-First Perspective

DGX agent

arXiv:2605.00505v1 Announce Type: cross Abstract: Modern information retrieval (IR) is no longer consumed primarily by humans but increasingly by large language models (LLMs) via retrieval-augmented g

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies

DGX agent

arXiv:2412.02125v2 Announce Type: replace-cross Abstract: Goal-conditioned policies enable decision-making models to execute diverse behaviors based on specified goals, yet their downstream performanc

model-releasesarxiv-cs-lg
4 May 2026
Local Ai

TimesNet-Gen: Deep Learning-based Site Specific Strong Motion Generation

DGX agent

arXiv:2512.04694v3 Announce Type: replace Abstract: Effective earthquake risk reduction relies on accurate site-specific evaluations, which require models capable of representing the influence of loca

local-aiarxiv-cs-lg
4 May 2026
Research

Understanding Cognitive States from Head & Hand Motion Data

DGX agent

arXiv:2509.24255v2 Announce Type: replace-cross Abstract: As virtual reality (VR) becomes widespread, head and hand motion data captured by consumer systems has become substantially more common. Howev

researcharxiv-cs-lg
4 May 2026
Research

Unlearning Offline Stochastic Multi-Armed Bandits

DGX agent

arXiv:2605.00638v1 Announce Type: new Abstract: Machine unlearning aims to unlearn data points from a learned model, offering a principled way to process data-deletion requests and mitigate privacy ri

researcharxiv-cs-lg
4 May 2026
Model Releases

VLAs are Confined yet Capable of Generalizing to Novel Instructions

DGX agent

arXiv:2505.03500v5 Announce Type: replace Abstract: Vision-language-action models (VLAs) often achieve high performance on demonstrated tasks but struggle significantly when required to extrapolate, c

model-releasesarxiv-cs-ro
4 May 2026
Model Releases

A Reproducibility Study of LLM-Based Query Reformulation

DGX agent

arXiv:2604.27421v1 Announce Type: cross Abstract: Large Language Models (LLMs) are now widely used for query reformulation and expansion in Information Retrieval, with many studies reporting substanti

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

AesRM: Improving Video Aesthetics with Expert-Level Feedback

DGX agent

arXiv:2604.28078v1 Announce Type: new Abstract: Despite rapid advances in photorealistic video generation, real-world applications such as filmmaking require video aesthetics, e.g., harmonious colors

model-releasesarxiv-cs-cv
1 May 2026
Safety

BicKD: Bilateral Contrastive Knowledge Distillation

DGX agent

arXiv:2602.01265v2 Announce Type: replace Abstract: Knowledge distillation (KD) is a machine learning framework that transfers knowledge from a teacher model to a student model. The vanilla KD propose

safetyarxiv-cs-lg
1 May 2026
Model Releases

COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts

DGX agent

arXiv:2604.27389v1 Announce Type: cross Abstract: In recent years, Multimodal Large Language Models (MLLMs) have achieved remarkable progress on a wide range of multimodal benchmarks. Despite these ad

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Decoding Scientific Experimental Images: The SPUR Benchmark for Perception, Understanding, and Reasoning

DGX agent

arXiv:2604.27604v1 Announce Type: new Abstract: We introduce SPUR, a comprehensive benchmark for scientific experimental image perception, understanding, and reasoning, comprising 4,264 question-answe

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

DeepTutor: Towards Agentic Personalized Tutoring

DGX agent

arXiv:2604.26962v1 Announce Type: cross Abstract: Education represents one of the most promising real-world applications for Large Language Models (LLMs). However, conventional tutoring systems rely o

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Diagnosing Capability Gaps in Fine-Tuning Data

DGX agent

arXiv:2604.27547v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) for domain-specific tasks requires training datasets that comprehensively cover the target capabilities a pract

model-releasesarxiv-cs-lg
1 May 2026
← Previous
1…453454455456457…1119
Next →