AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,405Total entries
1Added by human
92,404Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,927 results
Research

Mitigating Multimodal LLMs Hallucinations via Relevance Propagation at Inference Time

DGX agent

arXiv:2605.01766v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have revolutionized the landscape of AI, demonstrating impressive capabilities in tackling complex vision and

researcharxiv-cs-cv
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MOSAIC: Multi-agent Orchestration for Task-Intelligent Scientific Coding

DGX agent

arXiv:2510.08804v3 Announce Type: replace Abstract: We present MOSAIC, a multi-agent Large Language Model (LLM) framework for solving challenging scientific coding tasks. Unlike general-purpose coding

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety

DGX agent

arXiv:2605.01687v1 Announce Type: new Abstract: We present MultiBreak, a scalable and diverse multi-turn jailbreak benchmark to evaluate large language model (LLM) safety. Multi-turn jailbreaks mimic

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Pandora's Regret: A Proper Scoring Rule for Evaluating Sequential Search

DGX agent

arXiv:2605.01936v1 Announce Type: new Abstract: In sequential search, alternatives are tested until the true class is found. Standard proper scoring rules like log loss are local, ignoring the ranking

model-releasesarxiv-cs-lg
5 May 2026
Research

ParaRNN: An Interpretable and Parallelizable Recurrent Neural Network for Time-Dependent Data

DGX agent

arXiv:2605.02692v1 Announce Type: cross Abstract: The proliferation of large-scale and structurally complex data has spurred the integration of machine learning methods into statistical modeling. Recu

researcharxiv-cs-lg
5 May 2026
Model Releases

Physics-Informed Neural Learning for State Reconstruction and Parameter Identification in Coupled Greenhouse Climate Dynamics

DGX agent

arXiv:2605.02524v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have recently emerged as a promising framework for integrating data-driven learning with physical knowledge. In

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Prescriptive Scaling Laws for Data Constrained Training

DGX agent

arXiv:2605.01640v1 Announce Type: cross Abstract: Training compute is increasingly outpacing the availability of high-quality data. This shifts the central challenge from optimal compute allocation to

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Probe-Geometry Alignment: Erasing the Cross-Sequence Memorization Signature Below Chance

DGX agent

arXiv:2605.01699v1 Announce Type: new Abstract: Recent attacks show that behavioural unlearning of large language models leaves internal traces recoverable by adversarial probes. We characterise where

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Projection-Free Transformers via Gaussian Kernel Attention

DGX agent

arXiv:2605.02144v1 Announce Type: new Abstract: Self-attention in Transformers is typically implemented as softmax(QK^op/sqrt{d})V, where Q=XW_Q, K=XW_K, and V=XW_V are learned linear projections of t

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces

DGX agent

arXiv:2605.02801v1 Announce Type: new Abstract: As large language model (LLM) agents evolve from isolated tool users into coordinated teams, reinforcement learning (RL) must optimize not only individu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Rethinking Multi-Label Node Classification: Do Tuned Classic GNNs Suffice?

DGX agent

arXiv:2605.01403v1 Announce Type: new Abstract: Multi-label node classification (MLNC) has recently been addressed by increasingly complex label-aware designs that explicitly model node-label interact

model-releasesarxiv-cs-lg
5 May 2026
Research

Rhamba: Region-Aware Hybrid Attention-Mamba Framework for Self-Supervised Learning in Resting-State fMRI

DGX agent

arXiv:2605.01240v1 Announce Type: new Abstract: Self-supervised pretraining is promising for large-scale neuroimaging, yet the impact of region-aware masking and hybrid sequence modeling remains under

researcharxiv-cs-lg
5 May 2026
Agents

S^3-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data

DGX agent

arXiv:2605.01248v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has enabled newer capabilities in models, such as agentic tool-use for search. However, these models struggle

agentsarxiv-cs-lg
5 May 2026
Model Releases

SAMamba3D: adapting Segment Anything for generalizable 3D segmentation of multiphase pore-scale images

DGX agent

arXiv:2605.00916v1 Announce Type: new Abstract: Reliable segmentation of multiphase pore-scale X-ray images of rocks is necessary to quantify fluid saturation, connectivity, and interfacial geometry.

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

SemEval-2026 Task 7: Everyday Knowledge Across Diverse Languages and Cultures

DGX agent

arXiv:2605.02601v1 Announce Type: new Abstract: We present our shared task on evaluating the adaptability of LLMs and NLP systems across multiple languages and cultures. The task data consist of an ex

model-releasesarxiv-cs-cl
5 May 2026
Safety

Skills as Verifiable Artifacts: A Trust Schema and a Biconditional Correctness Criterion for Human-in-the-Loop Agent Runtimes

DGX agent

arXiv:2605.00424v1 Announce Type: cross Abstract: Agent skills -- structured packages of instructions, scripts, and references that augment a large language model (LLM) without modifying the model its

safetyarxiv-cs-ai
5 May 2026
Model Releases

STEP: Warm-Started Visuomotor Policies with Spatiotemporal Consistency Prediction

DGX agent

arXiv:2602.08245v2 Announce Type: replace Abstract: Diffusion policies have recently emerged as a powerful paradigm for visuomotor control in robotic manipulation due to their ability to model the dis

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

SwiftChannel: Algorithm-Hardware Co-Design for Deep Learning-Based 5G Channel Estimation

DGX agent

arXiv:2605.01931v1 Announce Type: cross Abstract: Channel estimation is crucial in 5G communication networks for optimizing transmission parameters and ensuring reliable, high-speed communication. How

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Task-Driven Subspace Decomposition for Knowledge Sharing and Isolation in LoRA-based Continual Learning

DGX agent

arXiv:2603.00191v2 Announce Type: replace-cross Abstract: Continual Learning (CL) requires models to sequentially adapt to new tasks without forgetting old knowledge. Recently, Low-Rank Adaptation (Lo

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Teaching LLMs Brazilian Healthcare: Injecting Knowledge from Official Clinical Guidelines

DGX agent

arXiv:2605.01077v1 Announce Type: new Abstract: Brazil's Unified Health System (SUS) relies on official clinical guidelines that define diagnostic criteria, treatments, dosages, and monitoring procedu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

The Good, the Bad, and the Sampled: a No-Regret Approach to Safe Online Classification

DGX agent

arXiv:2510.01020v2 Announce Type: replace Abstract: We study sequential testing for a binary disease outcome when risk follows an unknown logistic model. At each round, the decision maker may either p

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Towards Lightest Low-Light Image Enhancement Architecture for Mobile Devices

DGX agent

arXiv:2507.04277v2 Announce Type: replace Abstract: Real-time low-light image enhancement on mobile and embedded devices requires models that balance visual quality and computational efficiency. Exist

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

VILAS: A VLA-Integrated Low-cost Architecture with Soft Grasping for Robotic Manipulation

DGX agent

arXiv:2605.02037v1 Announce Type: new Abstract: We present VILAS, a fully low-cost, modular robotic manipulation platform designed to support end-to-end vision-language-action (VLA) policy learning an

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Borrowed Geometry: Computational Reuse of Frozen Text-Pretrained Transformer Weights Across Modalities

DGX agent

arXiv:2605.00333v1 Announce Type: cross Abstract: Frozen Gemma 4 31B weights pretrained exclusively on text tokens, unmodified, transfer across modality boundaries through a thin trainable interface.

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection

DGX agent

arXiv:2605.00630v1 Announce Type: new Abstract: The proliferation of advanced AI video synthesis techniques poses an unprecedented challenge to digital video authenticity. Existing AI-generated video

model-releasesarxiv-cs-cv
4 May 2026
Research

Confidence Estimation in Automatic Short Answer Grading with LLMs

DGX agent

arXiv:2605.00200v1 Announce Type: new Abstract: Automatic Short Answer Grading (ASAG) with generative large language models (LLMs) has recently demonstrated strong performance without task-specific fi

researcharxiv-cs-cl
4 May 2026
Model Releases

Deepfakes: we need to re-think the concept of 'real' images

DGX agent

arXiv:2509.21864v2 Announce Type: replace Abstract: The wide availability and low usability barrier of modern image generation models has triggered the reasonable fear of criminal misconduct and negat

model-releasesarxiv-cs-cv
4 May 2026
Research

End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer

DGX agent

arXiv:2605.00503v1 Announce Type: new Abstract: Autoregressive image modeling relies on visual tokenizers to compress images into compact latent representations. We design an end-to-end training pipel

researcharxiv-cs-cv
4 May 2026
Model Releases

FedACT: Concurrent Federated Intelligence across Heterogeneous Data Sources

DGX agent

arXiv:2605.00011v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative intelligence across decentralized data source devices in a privacy-preserving way. While substantial resea

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios

DGX agent

arXiv:2605.00706v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied in financial scenarios. However, they may produce harmful outputs, including facilitating illegal

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Firestore at Next '26: Unlock agentic development, search and MongoDB compatibility

DGX agent

In the era of AI agents, the distance between a big idea and a working application has never been shorter. As we lean more heavily on agents to help us build applications, a critical question remains:

model-releasesgoogle-cloud-ai
4 May 2026
Research

Flow matching for Sentinel-2 super-resolution: implementation, application, and implications

DGX agent

arXiv:2605.00367v1 Announce Type: new Abstract: Developing robust techniques for super-resolution of satellite imagery involves navigating commonly observed trade-offs between spectral fidelity and pe

researcharxiv-cs-cv
4 May 2026
Safety

FreeRet: MLLMs as Training-Free Retrievers

DGX agent

arXiv:2509.24621v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are emerging as versatile foundations for mixed-modality retrieval. Yet, they often require heavy post-hoc

safetyarxiv-cs-cv
4 May 2026
Model Releases

How OpenAI delivers low-latency voice AI at scale

DGX agent

OpenAI describes its technical approach to delivering real-time voice AI services with minimal latency across large user bases, likely covering infrastructure optimization, model serving strategies, a

model-releasesopenai
4 May 2026
Model Releases

Hypergraph and Latent ODE Learning for Multimodal Root Cause Localization in Microservices

DGX agent

arXiv:2605.00351v1 Announce Type: new Abstract: Root cause localization in cloud native microservice systems requires modeling complex service dependencies, irregular temporal dynamics, and heterogene

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Learning from the Unseen: Generative Data Augmentation for Geometric-Semantic Accident Anticipation

DGX agent

arXiv:2605.00051v1 Announce Type: new Abstract: Anticipating traffic accidents is a critical yet unresolved problem for autonomous driving, hindered by the inherent complexity of modeling interactions

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

LLM-Oriented Information Retrieval: A Denoising-First Perspective

DGX agent

arXiv:2605.00505v1 Announce Type: cross Abstract: Modern information retrieval (IR) is no longer consumed primarily by humans but increasingly by large language models (LLMs) via retrieval-augmented g

model-releasesarxiv-cs-cl
4 May 2026
Industry

Powering the Inference Era: Inside the DigitalOcean AI-Native Cloud

DGX agent

DigitalOcean outlines its AI-native cloud infrastructure designed to support the inference phase of AI model deployment, emphasizing tools and services that enable businesses to run and scale AI model

industrydigitalocean
4 May 2026
Model Releases

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies

DGX agent

arXiv:2412.02125v2 Announce Type: replace-cross Abstract: Goal-conditioned policies enable decision-making models to execute diverse behaviors based on specified goals, yet their downstream performanc

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

shipped Promptloop 🌀 Claude Code but for your prompts. a CLI agent that runs the full prompt-eval loop in one chat session: create test cas…

DGX agent

shipped Promptloop 🌀 Claude Code but for your prompts. a CLI agent that runs the full prompt-eval loop in one chat session: create test cases → run across models → generate reports → propose prompt →

model-releasesharrison-chase--x
4 May 2026
Research

The distillation panic

DGX agent

The article discusses concerns about the widespread use of knowledge distillation in AI development, where larger models' outputs are used to train smaller models, potentially creating a cycle of degr

researchinterconnects
4 May 2026
Local Ai

TimesNet-Gen: Deep Learning-based Site Specific Strong Motion Generation

DGX agent

arXiv:2512.04694v3 Announce Type: replace Abstract: Effective earthquake risk reduction relies on accurate site-specific evaluations, which require models capable of representing the influence of loca

local-aiarxiv-cs-lg
4 May 2026
Model Releases

Try Grok

DGX agent

Try Grok Grok 4.3 just became the smartest AI in the world at law and money It took #1 on TWO brutal private tests no other model could win on “Vals AI” benchmarks #1 CaseLaw (v2) - 79.31% accuracy Pr

model-releaseselon-musk--x
4 May 2026
Research

Understanding Cognitive States from Head & Hand Motion Data

DGX agent

arXiv:2509.24255v2 Announce Type: replace-cross Abstract: As virtual reality (VR) becomes widespread, head and hand motion data captured by consumer systems has become substantially more common. Howev

researcharxiv-cs-lg
4 May 2026
Research

Unlearning Offline Stochastic Multi-Armed Bandits

DGX agent

arXiv:2605.00638v1 Announce Type: new Abstract: Machine unlearning aims to unlearn data points from a learned model, offering a principled way to process data-deletion requests and mitigate privacy ri

researcharxiv-cs-lg
4 May 2026
Model Releases

VLAs are Confined yet Capable of Generalizing to Novel Instructions

DGX agent

arXiv:2505.03500v5 Announce Type: replace Abstract: Vision-language-action models (VLAs) often achieve high performance on demonstrated tasks but struggle significantly when required to extrapolate, c

model-releasesarxiv-cs-ro
4 May 2026
Model Releases

A Reproducibility Study of LLM-Based Query Reformulation

DGX agent

arXiv:2604.27421v1 Announce Type: cross Abstract: Large Language Models (LLMs) are now widely used for query reformulation and expansion in Information Retrieval, with many studies reporting substanti

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

AesRM: Improving Video Aesthetics with Expert-Level Feedback

DGX agent

arXiv:2604.28078v1 Announce Type: new Abstract: Despite rapid advances in photorealistic video generation, real-world applications such as filmmaking require video aesthetics, e.g., harmonious colors

model-releasesarxiv-cs-cv
1 May 2026
← Previous
1…556557558559560…1395
Next →