AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,553 results
5 May 2026

MolViBench: Evaluating LLMs on Molecular Vibe Coding

Model ReleasesDGX agent

arXiv:2605.02351v1 Announce Type: new Abstract: Molecular Vibe Coding, a paradigm where chemists interact with LLMs to generate executable programs for molecular tasks, has emerged as a flexible alter

MorphIt: Flexible Spherical Approximation of Robot Morphology for Representation-driven Adaptation

Model ReleasesDGX agent

arXiv:2507.14061v2 Announce Type: replace Abstract: What if a robot could rethink its own morphological representation to better meet the demands of diverse tasks? Most robotic systems today treat the

MOSAIC: Multi-agent Orchestration for Task-Intelligent Scientific Coding

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2510.08804v3 Announce Type: replace Abstract: We present MOSAIC, a multi-agent Large Language Model (LLM) framework for solving challenging scientific coding tasks. Unlike general-purpose coding

MPCS: Neuroplastic Continual Learning via Multi-Component Plasticity and Topology-Aware EWC

Model ReleasesDGX agent

arXiv:2605.02509v1 Announce Type: new Abstract: Continual learning systems face a fundamental tension between plasticity -- acquiring new knowledge -- and stability -- retaining prior knowledge. We in

Multi-fidelity surrogates for mechanics of composites: from co-kriging to multi-fidelity neural networks

Model ReleasesDGX agent

arXiv:2605.02871v1 Announce Type: cross Abstract: Composite materials exhibit strongly hierarchical and anisotropic properties governed by coupled mechanisms spanning constituents, plies, laminates, s

Multi-Perspective Transformers in ARC-AGI-2 Challenge

Model ReleasesDGX agent

arXiv:2605.01154v1 Announce Type: new Abstract: ARC-AGI-2 is a benchmark of human-intuitive visual puzzles that measures a machine's ability to generalize from limited examples, interpret symbolic mea

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety

Model ReleasesDGX agent

arXiv:2605.01687v1 Announce Type: new Abstract: We present MultiBreak, a scalable and diverse multi-turn jailbreak benchmark to evaluate large language model (LLM) safety. Multi-turn jailbreaks mimic

Multiple Choice Questions: Reasoning Makes Large Language Models (LLMs) More Self-Confident, Especially When They are Wrong

Model ReleasesDGX agent

arXiv:2501.09775v3 Announce Type: replace Abstract: Multiple Choice Question (MCQ) tests are among the most used methods for evaluating large language models (LLMs). Besides checking the correctness o

NAKUL-Med: Spectral-Graph State Space Models with Dynamics Kernels for Medical Signals

Model ReleasesDGX agent

arXiv:2605.00871v1 Announce Type: cross Abstract: State space models (SSMs) achieve linear-time complexity but struggle with multi-channel physiological signals due to three limitations: fixed kernels

Neighbor2Inverse: Self-Supervised Denoising for Low-Dose Region-of-Interest Phase Contrast CT

Model ReleasesDGX agent

arXiv:2605.01075v1 Announce Type: new Abstract: Propagation-based X-ray phase-contrast imaging (PBI) enables high-contrast visualization of lung structures and holds strong medical potential. However,

New ways to buy ChatGPT ads

Model ReleasesDGX agent

OpenAI introduced new advertising purchasing options for ChatGPT, expanding how businesses can buy ad placements within the platform. These new methods likely provide advertisers with additional flexi

Object-Level Explanations for Image Geolocation Models: a GeoGuessr use-case

Model ReleasesDGX agent

arXiv:2605.00912v1 Announce Type: new Abstract: When humans play geolocation games such as GeoGuessr, they rely on concrete visual cues, such as road markings, vegetation, or architectural details, to

OceanPile: A Large-Scale Multimodal Ocean Corpus for Foundation Models

Model ReleasesDGX agent

arXiv:2605.00877v1 Announce Type: cross Abstract: The vast and underexplored ocean plays a critical role in regulating global climate and supporting marine biodiversity, yet artificial intelligence ha

oMeBench: Towards Robust Benchmarking of LLMs in Organic Mechanism Elucidation and Reasoning

Model ReleasesDGX agent

arXiv:2510.07731v3 Announce Type: replace-cross Abstract: Organic reaction mechanisms are the stepwise elementary reactions by which reactants form intermediates and products, and are fundamental to u

Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection

Model ReleasesDGX agent

arXiv:2605.01638v1 Announce Type: new Abstract: Multimodal deepfakes are proliferating on social media and threaten authenticity, information integrity, and digital forensics. Existing benchmarks are

OmniTrack++: Omnidirectional Multi-Object Tracking by Learning Large-FoV Trajectory Feedback

Model ReleasesDGX agent

arXiv:2511.00510v2 Announce Type: replace Abstract: To address panoramic distortion, large search space, and identity ambiguity under a 360{eg} FoV, OmniTrack++ adopts a feedback-driven framework that

On Stable Long-Form Generation: Benchmarking and Mitigating Length Volatility

Model ReleasesDGX agent

arXiv:2605.01357v1 Announce Type: new Abstract: Large Language Models (LLMs) excel at long-context understanding but exhibit significant limitations in long-form generation. Existing studies primarily

On the Effectiveness of Integration Methods for Multimodal Dialogue Response Retrieval

Model ReleasesDGX agent

arXiv:2506.11499v2 Announce Type: replace Abstract: Multimodal chatbots have become one of the major topics for dialogue systems in both research community and industry. Recently, researchers have she

On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization

Model ReleasesDGX agent

arXiv:2605.02141v1 Announce Type: new Abstract: Kullback-Leibler (KL) regularization is widely used in offline decision-making and offers several benefits, motivating recent work on the sample complex

Open models should compete on cost and specialization, not frontier benchmarks @natolambert puts it well: the right benchmark is savings in …

Model ReleasesDGX agent

Open models should compete on cost and specialization, not frontier benchmarks @natolambert puts it well: the right benchmark is savings in compute and time, especially for repetitive agent tasks deep

OpenAI claims ChatGPT’s new default model hallucinates way less

Model ReleasesDGX agent

OpenAI's newest default model for ChatGPT might not make stuff up as much. Hallucinations have been an ongoing problem for AI models, but OpenAI says its new GPT-5.5 Instant model has 'significant imp

OpenAI GPT-5 System Card

Model ReleasesDGX agent

arXiv:2601.03267v2 Announce Type: replace Abstract: This is the system card published alongside the OpenAI GPT-5 launch, August 2025. GPT-5 is a unified system with a smart and fast model that answers

OphMAE: Bridging Volumetric and Planar Imaging with a Foundation Model for Adaptive Ophthalmological Diagnosis

Model ReleasesDGX agent

arXiv:2605.02714v1 Announce Type: new Abstract: The advent of foundation models has heralded a new era in medical artificial intelligence (AI), enabling the extraction of generalizable representations

OralMLLM-Bench: Evaluating Cognitive Capabilities of Multimodal Large Language Models in Dental Practice

Model ReleasesDGX agent

arXiv:2605.01333v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have emerged as a promising paradigm for dental image analysis. However, their ability to capture the multi-lev

Orchestrating Spatial Semantics via a Zone-Graph Paradigm for Intricate Indoor Scene Generation

Model ReleasesDGX agent

arXiv:2605.02537v1 Announce Type: new Abstract: Autonomous 3D indoor scene synthesis breaks down in non-convex rooms with tightly coupled spatial constraints. Data-driven generators lack topological p

Orthographic Constraint Satisfaction and Human Difficulty Alignment in Large Language Models

Model ReleasesDGX agent

arXiv:2511.21086v2 Announce Type: replace Abstract: Large language models must satisfy hard orthographic constraints during controlled text generation, yet systematic cross-family evaluation remains l

PACE: Parameter Change for Unsupervised Environment Design

Model ReleasesDGX agent

arXiv:2605.01358v1 Announce Type: new Abstract: Unsupervised Environment Design (UED) offers a promising paradigm for improving reinforcement learning generalization by adaptively shaping training env

Pair2Score: Pairwise-to-Absolute Transfer for LLM-Based Essay Scoring

Model ReleasesDGX agent

arXiv:2605.02069v1 Announce Type: new Abstract: Many scoring applications require absolute predictions, while pairwise comparisons can provide a simpler learning objective. We present Pair2Score, a tw

Pandora's Regret: A Proper Scoring Rule for Evaluating Sequential Search

Model ReleasesDGX agent

arXiv:2605.01936v1 Announce Type: new Abstract: In sequential search, alternatives are tested until the true class is found. Standard proper scoring rules like log loss are local, ignoring the ranking

Parameter Space Analysis through Guided Visual Interpolations

Model ReleasesDGX agent

arXiv:2509.19202v2 Announce Type: replace-cross Abstract: We propose Parameter Space Analysis through Guided Visual Interpolations (ParamInter), a novel tool for high-dimensional input parameter space

PC-MNet: Dual-Level Congruity Modeling for Multimodal Sarcasm Detection via Polarity-Modulated Attention

Model ReleasesDGX agent

arXiv:2605.02447v1 Announce Type: new Abstract: Multimodal sarcasm detection, which aims to precisely identify pragmatic incongruities between literal text and nonverbal cues, has gained substantial a

PepSpecBench: A Unified Evaluation Benchmark for Peptide Tandem Mass Spectrometry Prediction

Model ReleasesDGX agent

arXiv:2605.01945v1 Announce Type: new Abstract: Tandem mass spectrometry provides a high-throughput framework for identifying and quantifying proteins in complex biological samples. In computational p

Perturbation Dose Responses in Recursive LLM Loops: Raw Switching, Stochastic Floors, and Persistent Escape under Append, Replace, and Dialog Updates

Model ReleasesDGX agent

arXiv:2605.02236v1 Announce Type: cross Abstract: Recursive language-model loops often settle into recognizable attractor-like patterns. The practical question is how much injected text is needed to m

PhaseNet++: Phase-Aware Frequency-Domain Anomaly Detection for Industrial Control Systems via Phase Coherence Graphs

Model ReleasesDGX agent

arXiv:2605.00929v1 Announce Type: new Abstract: Multivariate time series anomaly detection in ICS has attracted growing attention due to the increasing threat of cyber-physical attacks on critical inf

Physics-Informed Neural Learning for State Reconstruction and Parameter Identification in Coupled Greenhouse Climate Dynamics

Model ReleasesDGX agent

arXiv:2605.02524v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have recently emerged as a promising framework for integrating data-driven learning with physical knowledge. In

Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning

Model ReleasesDGX agent

arXiv:2605.02168v1 Announce Type: cross Abstract: Language model (LM)-based agents have demonstrated promising capabilities in automating complex tasks from natural language instructions, yet they con

PointCSP: Cross-Sample Semantic Propagation and Stability Preservation in Self-Supervised Point Cloud Learning

Model ReleasesDGX agent

arXiv:2605.01759v1 Announce Type: new Abstract: Scene-level point cloud self-supervised learning (PC-SSL) has demonstrated potential in enhancing the generalization capability of 3D vision models. Des

PPO guided Agentic Pipeline for Adaptive Prompt Selection and Test Case Generation

Model ReleasesDGX agent

arXiv:2605.00942v1 Announce Type: cross Abstract: Developing effective test cases capable of thoroughly exercising large-scale software systems is inherently difficult, especially if such systems have

Prescriptive Scaling Laws for Data Constrained Training

Model ReleasesDGX agent

arXiv:2605.01640v1 Announce Type: cross Abstract: Training compute is increasingly outpacing the availability of high-quality data. This shifts the central challenge from optimal compute allocation to

Pretraining A Large Language Model using Distributed GPUs: A Memory-Efficient Decentralized Paradigm

Model ReleasesDGX agent

arXiv:2602.11543v2 Announce Type: replace Abstract: Pretraining large language models (LLMs) typically requires centralized clusters with thousands of high-memory GPUs (e.g., H100/A100). Recent decent

PRIME: Protein Representation via Physics-Informed Multiscale Equivariant Hierarchies

Model ReleasesDGX agent

arXiv:2605.01625v1 Announce Type: new Abstract: Proteins are inherently multiscale physical systems whose functional properties emerge from coordinated structural organization across multiple spatial

Probe-Geometry Alignment: Erasing the Cross-Sequence Memorization Signature Below Chance

Model ReleasesDGX agent

arXiv:2605.01699v1 Announce Type: new Abstract: Recent attacks show that behavioural unlearning of large language models leaves internal traces recoverable by adversarial probes. We characterise where

Projection-Free Transformers via Gaussian Kernel Attention

Model ReleasesDGX agent

arXiv:2605.02144v1 Announce Type: new Abstract: Self-attention in Transformers is typically implemented as softmax(QK^op/sqrt{d})V, where Q=XW_Q, K=XW_K, and V=XW_V are learned linear projections of t

Prosa: Rubric-Based Evaluation of LLMs on Real User Chats in Brazilian Portuguese

Model ReleasesDGX agent

arXiv:2605.01630v1 Announce Type: new Abstract: Rankings produced by holistic LLM-as-a-judge scoring are sensitive to the bias of the chosen judge model. We show that switching to binary rubric scorin

Proud of my analyst, Hermes ​ This is what he has achieved after 1.5+ months of internship with me ​ - Got me off X (doomscrolling cut from …

Model ReleasesDGX agent

Proud of my analyst, Hermes ​ This is what he has achieved after 1.5+ months of internship with me ​ - Got me off X (doomscrolling cut from 2 hrs → 30 mins) - Practically paid for his salary via PMs b

Psychologically Potent, Computationally Invisible: LLMs Generate Social-Comparison Triggers They Fail to Detect

Model ReleasesDGX agent

arXiv:2605.01017v1 Announce Type: new Abstract: We introduce Xiaohongshu Social Comparison Reader Elicitation (XHS-SCoRE), a reader-grounded benchmark for detecting if a text-only Xiaohongshu (RedNote

Public sector momentum and mission impact at Google Cloud Next ‘26

Model ReleasesDGX agent

The agentic era is here, and the public sector is at the forefront of leading this transformation.At Google Cloud Next ’26, it was clear that academia and public sector organizations are no longer jus

Quaternion Nonlinear Transform-Induced Nuclear Norm for Low-Rank Tensor Completion

Model ReleasesDGX agent

arXiv:2605.01467v1 Announce Type: cross Abstract: Tensor completion has emerged as a powerful framework for recovering missing data in multidimensional signals by exploiting low-rank tensor structures

RADMI: Latent Information Aggregation as a Proxy for Model Uncertainty

Model ReleasesDGX agent

arXiv:2605.01502v1 Announce Type: new Abstract: Epistemic uncertainty estimation is essential for identifying regions where deep learning system outputs may be unreliable. However, existing approaches

RAFNet: Region-Aware Fusion Network for Pansharpening

Model ReleasesDGX agent

arXiv:2605.02184v1 Announce Type: new Abstract: Pansharpening aims to generate high-resolution multispectral (HRMS) images by fusing low-resolution multispectral (LRMS) and high-resolution panchromati

RamanBench: A Large-Scale Benchmark for Machine Learning on Raman Spectroscopy

Model ReleasesDGX agent

arXiv:2605.02003v1 Announce Type: new Abstract: Machine Learning (ML) has transformed many scientific fields, yet key applications still lack standardized benchmarks. Raman spectroscopy, a widely used

Real-Time Text Transmission via LLM-Based Entropy Coding over Fixed-Rate Channels

Model ReleasesDGX agent

arXiv:2605.01991v1 Announce Type: cross Abstract: Learning, prediction, and compression are intimately connected: a model that accurately predicts the next symbol in a sequence can be coupled with a s

Recurrent Deep Reinforcement Learning for Chemotherapy Control under Partial Observability

Model ReleasesDGX agent

arXiv:2605.02552v1 Announce Type: new Abstract: Chemotherapy dose optimization can be formulated as a dynamic treatment regime, requiring sequential decisions under uncertainty that must balance tumor

Reference-Sampled Boltzmann Projection for KL-Regularized RLVR: Target-Matched Weighted SFT, Finite One-Shot Gaps, and Policy Mirror Descent

Model ReleasesDGX agent

arXiv:2605.02469v1 Announce Type: new Abstract: Online reinforcement learning with verifiable rewards (RLVR) turns checkable outcomes into a scalable training signal, but it keeps rollout generation,

RefusalGuard: Geometry-Preserving Fine-Tuning for Safety in LLMs

Model ReleasesDGX agent

arXiv:2605.01913v1 Announce Type: cross Abstract: Fine-tuning safety-aligned language models for downstream tasks often leads to substantial degradation of refusal behavior, making models vulnerable t

Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces

Model ReleasesDGX agent

arXiv:2605.02801v1 Announce Type: new Abstract: As large language model (LLM) agents evolve from isolated tool users into coordinated teams, reinforcement learning (RL) must optimize not only individu

Rethinking Multi-Label Node Classification: Do Tuned Classic GNNs Suffice?

Model ReleasesDGX agent

arXiv:2605.01403v1 Announce Type: new Abstract: Multi-label node classification (MLNC) has recently been addressed by increasingly complex label-aware designs that explicitly model node-label interact

Retrieving Any Relevant Moments: Benchmark and Models for Generalized Moment Retrieval

Model ReleasesDGX agent

arXiv:2605.02623v1 Announce Type: new Abstract: Video Moment Retrieval (VMR) aims to localize temporal segments in videos that correspond to a natural language query, but typically assumes only a sing

RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences

Model ReleasesDGX agent

arXiv:2605.01831v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback has become the standard paradigm for language model alignment, where reward models directly determine alignme

Robust Adaptive Predictive Control for Hook-Based Aerial Transportation Between Moving Platforms

Model ReleasesDGX agent

arXiv:2605.02370v1 Announce Type: new Abstract: This paper presents a novel model predictive control (MPC) approach for autonomous pick-and-place between moving platforms with a hook-equipped aerial m

← Previous
1…288289290291292…376
Next →