AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,624 results
13 May 2026

GKnow: Measuring the Entanglement of Gender Bias and Factual Gender

Model ReleasesDGX agent

arXiv:2605.12299v1 Announce Type: new Abstract: Recent works have analyzed the impact of individual components of neural networks on gendered predictions, often with a focus on mitigating gender bias.

gym-invmgmt: An Open Benchmarking Framework for Inventory Management Methods

Model ReleasesDGX agent

arXiv:2605.11355v1 Announce Type: new Abstract: Inventory-policy comparisons are often difficult to interpret because performance depends on the evaluation contract as much as on the policy itself. Di

Hi-GaTA: Hierarchical Gated Temporal Aggregation Adapter for Surgical Video Report Generation

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.11208v1 Announce Type: new Abstract: Automated, clinician-grade assessment reports for surgical procedures could reduce documentation burden and provide objective feedback, yet remain chall

Images in Sentences: Scaling Interleaved Instructions for Unified Visual Generation

SafetyDGX agent

arXiv:2605.12305v1 Announce Type: new Abstract: While recent advancements in multimodal language models have enabled image generation from expressive multi-image instructions, existing methods struggl

Interpretable EEG Microstate Discovery via Variational Deep Embedding: A Systematic Architecture Search with Multi-Quadrant Evaluation

ResearchDGX agent

arXiv:2605.10947v1 Announce Type: new Abstract: EEG microstate analysis segments continuous brain electrical activity into brief, quasi-stable topographic configurations that reflect discrete function

// δ-mem: Efficient Online Memory for LLMs // One of the more elegant memory mechanisms I've seen this month. Most long-term memory work eit…

TutorialsDGX agent

// δ-mem: Efficient Online Memory for LLMs // One of the more elegant memory mechanisms I've seen this month. Most long-term memory work either inflates context or retrains the model. This paper shows

MobileEgo Anywhere: Open Infrastructure for long horizon egocentric data on commodity hardware

ResearchDGX agent

arXiv:2605.05945v2 Announce Type: replace-cross Abstract: The recent advancement of Vision Language Action (VLA) models has driven a critical demand for large scale egocentric datasets. However, exist

Multi-modal Bayesian Neural Network Surrogates with Conjugate Last-Layer Estimation

TutorialsDGX agent

arXiv:2509.21711v2 Announce Type: replace-cross Abstract: As data collection and simulation capabilities advance, multi-modal learning, the task of learning from multiple modalities and sources of dat

Multi-Timescale Conductance Spiking Networks: A Sparse, Gradient-Trainable Framework with Rich Firing Dynamics for Enhanced Temporal Processing

ResearchDGX agent

arXiv:2605.11835v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) promise low-power event-driven computation for temporally rich tasks, but commonly used neuron models often trade off g

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue

SafetyDGX agent

arXiv:2605.05630v2 Announce Type: replace Abstract: Hidden malicious intent in multi-turn dialogue poses a growing threat to deployed large language models (LLMs). Rather than exposing a harmful objec

Output Composability of QLoRA PEFT Modules for Plug-and-Play Attribute-Controlled Text Generation

Model ReleasesDGX agent

arXiv:2605.12345v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) techniques offer task-specific fine-tuning at a fraction of the cost of full fine-tuning, but require separate fi

Overview of the MedHopQA track at BioCreative IX: track description, participation and evaluation of systems for multi-hop medical question answering

Model ReleasesDGX agent

arXiv:2605.12313v1 Announce Type: new Abstract: Multi-hop question answering (QA) remains a significant challenge in the biomedical domain, requiring systems to integrate information across multiple s

Principled Latent Diffusion for Graphs via Laplacian Autoencoders

ResearchDGX agent

arXiv:2601.13780v3 Announce Type: replace Abstract: Graph diffusion models achieve state-of-the-art performance in graph generation but suffer from quadratic complexity in the number of nodes -- and m

PRISM: A Geometric Risk Bound that Decomposes Drift into Scale, Shape, and Head

Local AiDGX agent

arXiv:2605.11608v1 Announce Type: new Abstract: Comparing post-training LLM variants, such as quantized, LoRA-adapted, and distilled models, requires a diagnostic that identifies how a variant has dri

PrivacySIM: Evaluating LLM Simulation of User Privacy Behavior

ApplicationsDGX agent

arXiv:2605.12147v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to simulate human behavior, but their ability to simulate individual privacy decisions is not well

QDSB: Quantized Diffusion Schrodinger Bridges

ApplicationsDGX agent

arXiv:2605.11983v1 Announce Type: new Abstract: Learning generative models in settings where the source and target distributions are only specified through unpaired samples is gaining in importance. H

Quite excited about llama-eval, a proposed eval tool for llama.cpp. Could be a nice step toward more comparable community evals 🎉 https://g…

Model ReleasesDGX agent

Llama-eval is a proposed evaluation tool for llama.cpp designed to standardize and improve comparability of community-run evaluations. The tool aims to address inconsistencies in how different users b

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch

Model ReleasesDGX agent

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal i

Reflect then Learn: Active Prompting for Information Extraction Guided by Introspective Confusion

TutorialsDGX agent

arXiv:2508.10036v2 Announce Type: replace Abstract: Large Language Models (LLMs) show remarkable potential for few-shot information extraction (IE), yet their performance is highly sensitive to the ch

SafeManip: A Property-Driven Benchmark for Temporal Safety Evaluation in Robotic Manipulation

Model ReleasesDGX agent

arXiv:2605.12386v1 Announce Type: new Abstract: Robotic manipulation is typically evaluated by task success, but successful completion does not guarantee safe execution. Many safety failures are tempo

SAGAS: Semantic-Aware Graph-Assisted Stitching for Offline Temporal Logic Planning

SafetyDGX agent

arXiv:2512.00775v2 Announce Type: replace Abstract: Linear Temporal Logic (LTL) provides a rigorous framework for specifying long-horizon robotic tasks, yet existing approaches face a trade-off: model

Sharpen Your Flow: Sharpness-Aware Sampling for Flow Matching

TutorialsDGX agent

arXiv:2605.11547v1 Announce Type: new Abstract: Flow matching models generate samples by numerically integrating a learned velocity field, with each integration step requiring a neural network evaluat

Sparse Attention Remapping with Clustering for Efficient LLM Decoding on PIM

ResearchDGX agent

arXiv:2505.05772v2 Announce Type: replace Abstract: Transformer-based models are the foundation of modern machine learning, but their execution, particularly during autoregressive decoding in large la

Spatial Adapter: Structured Spatial Decomposition and Closed-Form Covariance for Frozen Predictors

Model ReleasesDGX agent

arXiv:2605.11394v1 Announce Type: cross Abstract: We present the Spatial Adapter, a parameter-efficient post-hoc layer that equips any frozen first-stage predictor with a structured spatial representa

SyncDPO: Enhancing Temporal Synchronization in Video-Audio Joint Generation via Preference Learning

Model ReleasesDGX agent

arXiv:2605.12179v1 Announce Type: new Abstract: Recent advancements in video-audio joint generation have achieved remarkable success in semantic correspondence. However, achieving precise temporal syn

The future of AI is agent-native. Excited to kick off this journey together with Hermes Agent and the @NousResearch community. Qwen 3.6 Plus…

Model ReleasesDGX agent

The future of AI is agent-native. Excited to kick off this journey together with Hermes Agent and the @NousResearch community. Qwen 3.6 Plus is now FREE for a limited time on Nous Portal — give it a t

The Scaling Law of Evaluation Failure: Why Simple Averaging Collapses Under Data Sparsity and Item Difficulty Gaps, and How Item Response Theory Recovers Ground Truth Across Domains

Model ReleasesDGX agent

arXiv:2605.11205v1 Announce Type: new Abstract: Benchmark evaluation across AI and safety-critical domains overwhelmingly relies on simple averaging. We demonstrate that this practice produces substan

Today we release Token Superposition Training (TST), a modification to the standard LLM pretraining loop that produces a 2-3× wall-clock spe…

ResearchDGX agent

Today we release Token Superposition Training (TST), a modification to the standard LLM pretraining loop that produces a 2-3× wall-clock speedup at matched FLOPs without changing the model architectur

TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing

Model ReleasesDGX agent

arXiv:2605.11473v1 Announce Type: cross Abstract: Soft Actor-Critic (SAC) and its variants dominate Multi-Task Reinforcement Learning (MTRL) due to their off-policy sample efficiency, while on-policy

trying more serious TNG content with LTX2.3

Local AiDGX agent

A user experiments with serious Star Trek: The Next Generation styled video content using LTX-2.3, a video generation model with available LoRA style adapters. LTX-2.3 is a major upgrade to video gene

🎉 We published a new AI safety study: shopping agents fall for whimsical attacks and lose money. A whimsical attack is an absurd scenario a…

Model ReleasesDGX agent

🎉 We published a new AI safety study: shopping agents fall for whimsical attacks and lose money. A whimsical attack is an absurd scenario a human would never try on another human. In one run, GPT-5.1

What Does It Mean for a Medical AI System to Be Right?

Model ReleasesDGX agent

arXiv:2605.11963v1 Announce Type: new Abstract: This paper examines what it means for a medical AI system to be right by grounding the question in a specific clinical context: the automatic classifica

12 May 2026

3DReflecNet: A Large-Scale Dataset for 3D Reconstruction of Reflective, Transparent, and Low-Texture Objects

Model ReleasesDGX agent

arXiv:2605.10204v1 Announce Type: new Abstract: Accurate 3D reconstruction of objects with reflective, transparent, or low-texture surfaces still remains notoriously challenging. Such materials often

A Quantum Inspired Variational Kernel and Explainable AI Framework for Cross Region Solar and Wind Energy Forecasting

Model ReleasesDGX agent

arXiv:2605.09032v1 Announce Type: cross Abstract: Reliable short horizon forecasting of solar and wind generation is a structural prerequisite of any modern power system yet most published forecasters

A Reflective Storytelling Agent for Older Adults: Integrating Argumentation Schemes and Argument Mining in LLM-Based Personalised Narratives

AgentsDGX agent

arXiv:2605.10531v1 Announce Type: new Abstract: This work investigates whether knowledge-driven large language model (LLM)-based storytelling can support purposeful narrative interaction with a digita

A Semantic-Sampling Framework for Evaluating Calibration in Open-Ended Question Answering

TutorialsDGX agent

arXiv:2605.08432v1 Announce Type: cross Abstract: Calibration measures whether a model's predicted confidence aligns with its empirical accuracy, and is central to the reliable deployment of large lan

A Spectral Framework for Closed-Form Relative Density Estimation

ResearchDGX agent

arXiv:2605.10668v1 Announce Type: new Abstract: We propose a closed-form spectral framework for relative log-density estimation in linearly parameterized probabilistic models, including unnormalized a

A Stability Benchmark of Generative Regularizers for Inverse Problems

Model ReleasesDGX agent

arXiv:2605.10076v1 Announce Type: cross Abstract: Generative (diffusion) priors demonstrate remarkable performance in addressing inverse problems in imaging. Yet, for scientific and medical imaging, i

AAAC: Activation-Aware Adaptive Codebooks for 4-bit LLM Weight Quantization

HardwareDGX agent

arXiv:2605.08692v1 Announce Type: cross Abstract: Post-training weight-only quantization to 4 bits is widely used to reduce the memory and compute costs of large language model inference. Existing PTQ

Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities

ApplicationsDGX agent

arXiv:2605.09678v1 Announce Type: new Abstract: While extremely powerful and versatile at various tasks, the thinking capabilities of large language models (LLMs) are often put under scrutiny as they

Acceptance Cards:A Four-Diagnostic Standard for Safe Fine-Tuning Defense Claims

Model ReleasesDGX agent

arXiv:2605.10575v1 Announce Type: cross Abstract: Safe fine-tuning defenses are often endorsed on the basis of a held-out gap reduction, but the same reduction can come from sampling noise, subject ar

AdaPreLoRA: Adafactor Preconditioned Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2605.08734v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) reparameterizes a weight update as a product of two low-rank factors, but the Jacobian J_{G} of the generator mapping the f

Adaptive digital twins for predictive decision-making: Online Bayesian learning of transition dynamics

ApplicationsDGX agent

arXiv:2512.13919v2 Announce Type: replace Abstract: This work shows how adaptivity can enhance value realization of digital twins in civil engineering. We focus on adapting the state transition models

Adaptive Multi-view Graph Contrastive Learning via Fractional-order Neural Diffusion Networks

Model ReleasesDGX agent

arXiv:2511.06216v4 Announce Type: replace Abstract: Graph contrastive learning (GCL) learns node and graph representations by contrasting multiple views of the same graph. Existing methods typically r

AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.08715v1 Announce Type: cross Abstract: LLM-based multi-agent systems are increasingly deployed on long-horizon tasks, but a single decisive error is often accepted by downstream agents and

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

Model ReleasesDGX agent

arXiv:2605.10286v1 Announce Type: new Abstract: Building effective clinical decision support systems requires the synthesis of complex heterogeneous multimodal data. Such modalities include temporal e

AI has a secondary intent problem. I was negotiating a rent renewal (I live in NYC and they raised rent 10% because I guess they were bored)…

Model ReleasesDGX agent

AI has a secondary intent problem. I was negotiating a rent renewal (I live in NYC and they raised rent 10% because I guess they were bored). The property manager said she would 'do everything she cou

Attention-Mamba: A Mamba-Enhanced Multi-Scale Parallel Inference Network for Medical Image Segmentation

SafetyDGX agent

arXiv:2402.02286v4 Announce Type: replace-cross Abstract: U-shaped architectures have long dominated the field of medical image segmentation, while Transformers are widely employed for modeling long-r

AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs

ResearchDGX agent

arXiv:2509.08031v3 Announce Type: replace-cross Abstract: Large Audio Language Models (LALMs) are rapidly advancing, but evaluating them remains challenging due to inefficient and non-standardized too

Auto-Rubric as Reward: From Implicit Preferences to Explicit Multimodal Generative Criteria

SafetyDGX agent

arXiv:2605.08354v1 Announce Type: new Abstract: Aligning multimodal generative models with human preferences demands reward signals that respect the compositional, multi-dimensional structure of human

Automated Detection of Abnormalities in Zebrafish Development

ResearchDGX agent

arXiv:2605.10464v1 Announce Type: new Abstract: Zebrafish embryos are a valuable model for drug discovery due to their optical transparency and genetic similarity to humans. However, current evaluatio

BEACON: A Multimodal Dataset for Learning Behavioral Fingerprints from Gameplay Data

Model ReleasesDGX agent

arXiv:2605.10867v1 Announce Type: cross Abstract: Continuous authentication in high-stakes digital environments requires datasets with fine-grained behavioral signals under realistic cognitive and mot

BenchHAR: Benchmarking Self-Supervised Learning for Generalizable Sensor-based Activity Recognition

Model ReleasesDGX agent

arXiv:2605.08296v1 Announce Type: new Abstract: Human Activity Recognition (HAR) from wearable sensors supports broad healthcare and behavior science applications. However, data heterogeneity and the

Beyond Toy Benchmarks: A Systematic Evaluation of OOD Detection Methods For Plant Pathology Classification

Model ReleasesDGX agent

arXiv:2605.08618v1 Announce Type: new Abstract: Out-of-distribution (OOD) detection is essential for reliable deployment of deep learning systems, yet the majority of existing methods are evaluated on

Black-Box Detection of LLM-Generated Text Using Generalized Jensen-Shannon Divergence

ResearchDGX agent

arXiv:2510.07500v2 Announce Type: replace Abstract: We study black-box detection of machine-generated text under practical constraints: the scoring model (proxy LM) may mismatch the unknown source mod

Breaking Contextual Inertia: Reinforcement Learning with Single-Turn Anchors for Stable Multi-Turn Interaction

ResearchDGX agent

arXiv:2603.04783v2 Announce Type: replace Abstract: While LLMs demonstrate strong reasoning capabilities when provided with full information in a single turn, they exhibit substantial vulnerability in

Byte-Exact Deduplication in Retrieval-Augmented Generation: A Three-Regime Empirical Analysis Across Public Benchmarks

Model ReleasesDGX agent

arXiv:2605.09611v1 Announce Type: new Abstract: This preprint presents an empirical analysis of byte-exact chunk-level deduplication in Retrieval-Augmented Generation (RAG) pipelines. We measure conte

CellDX AI Autopilot: Agent-Guided Training and Deployment of Pathology Classifiers

HardwareDGX agent

arXiv:2605.10362v1 Announce Type: new Abstract: Training AI models for computational pathology currently requires access to expensive whole-slide-image datasets, GPU infrastructure, deep expertise in

CNSocialDepress: A Chinese Social Media Dataset for Depression Risk Detection and Structured Analysis

Model ReleasesDGX agent

arXiv:2510.11233v3 Announce Type: replace Abstract: Depression is a pressing global public health issue, yet publicly available Chinese-language resources for depression risk detection remain scarce a

Combining Mechanical and Agentic Specification Inference for Move

Model ReleasesDGX agent

arXiv:2605.10005v1 Announce Type: cross Abstract: In this paper, we describe early work on a specification inference tool for the Move Prover that combines a weakest-precondition (WP) analysis over Mo

← Previous
1…560561562563564…1061
Next →