AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
24 Apr 2026

Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs

Model ReleasesDGX agent

arXiv:2604.20945v1 Announce Type: cross Abstract: Effective safety auditing of large language models (LLMs) demands tools that go beyond black-box probing and systematically uncover vulnerabilities ro

Clinically-Informed Modeling for Pediatric Brain Tumor Classification from Whole-Slide Histopathology Images

ResearchDGX agent

arXiv:2604.21060v1 Announce Type: new Abstract: Accurate diagnosis of pediatric brain tumors, starting with histopathology, presents unique challenges for deep learning, including severe data scarcity

GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR

Model ReleasesDGX agent

arXiv:2601.09361v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a key paradigm for improving large-scale reasoning models. Unlike supervised fine-tun

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

How VLAs (Really) Work In Open-World Environments

SafetyDGX agent

arXiv:2604.21192v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have been extensively used in robotics applications, achieving great success in various manipulation problems. Mo

Hyperloop Transformers

Model ReleasesDGX agent

arXiv:2604.21254v1 Announce Type: cross Abstract: LLM architecture research generally aims to maximize model quality subject to fixed compute/latency budgets. However, many applications of interest su

Not-a-Bandit: Provably No-Regret Drafter Selection in Speculative Decoding for LLMs

ResearchDGX agent

arXiv:2510.20064v2 Announce Type: replace Abstract: Speculative decoding is widely used in accelerating large language model (LLM) inference. In this work, we focus on the online draft model selection

OpenAI GPT-5.5 now available on Databricks, fully-governed through Unity AI Gateway

Model ReleasesDGX agent

OpenAI's GPT-5.5 model is now available on Databricks' platform with governance capabilities provided through Unity AI Gateway, enabling enterprises to deploy and manage the model within their data in

OpenEstimate: Evaluating LLMs on Reasoning Under Uncertainty with Real-World Data

Model ReleasesDGX agent

arXiv:2510.15096v2 Announce Type: replace Abstract: Real-world settings where language models (LMs) are deployed -- in domains spanning healthcare, finance, and other forms of knowledge work -- requir

Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies

Model ReleasesDGX agent

arXiv:2604.21571v1 Announce Type: new Abstract: Current model training approaches incorporate user information directly into shared weights, making individual data removal computationally infeasible w

Towards Universal Tabular Embeddings: A Benchmark Across Data Tasks

Model ReleasesDGX agent

arXiv:2604.21696v1 Announce Type: new Abstract: Tabular foundation models aim to learn universal representations of tabular data that transfer across tasks and domains, enabling applications such as t

VARestorer: One-Step VAR Distillation for Real-World Image Super-Resolution

Model ReleasesDGX agent

arXiv:2604.21450v1 Announce Type: cross Abstract: Recent advancements in visual autoregressive models (VAR) have demonstrated their effectiveness in image generation, highlighting their potential for

VistaBot: View-Robust Robot Manipulation via Spatiotemporal-Aware View Synthesis

Model ReleasesDGX agent

arXiv:2604.21914v1 Announce Type: new Abstract: Recently, end-to-end robotic manipulation models have gained significant attention for their generalizability and scalability. However, they often suffe

Wiring the 'Why': A Unified Taxonomy and Survey of Abductive Reasoning in LLMs

Model ReleasesDGX agent

arXiv:2604.08016v2 Announce Type: replace Abstract: Regardless of its foundational role in human discovery and sense-making, abductive reasoning--the inference of the most plausible explanation for an

23 Apr 2026

AFMRL: Attribute-Enhanced Fine-Grained Multi-Modal Representation Learning in E-commerce

ResearchDGX agent

arXiv:2604.20135v1 Announce Type: new Abstract: Multimodal representation is crucial for E-commerce tasks such as identical product retrieval. Large representation models (e.g., VLM2Vec) demonstrate s

CCTVBench: Contrastive Consistency Traffic VideoQA Benchmark for Multimodal LLMs

Model ReleasesDGX agent

arXiv:2604.20460v1 Announce Type: new Abstract: Safety-critical traffic reasoning requires contrastive consistency: models must detect true hazards when an accident occurs, and reliably reject plausib

DAIRE: A lightweight AI model for real-time detection of Controller Area Network attacks in the Internet of Vehicles

SafetyDGX agent

arXiv:2604.20771v1 Announce Type: cross Abstract: The Internet of Vehicles (IoV) is advancing modern transportation by improving safety, efficiency, and intelligence. However, the reliance on the Cont

Do Hallucination Neurons Generalize? Evidence from Cross-Domain Transfer in LLMs

ApplicationsDGX agent

arXiv:2604.19765v1 Announce Type: cross Abstract: Recent work identifies a sparse set of 'hallucination neurons' (H-neurons), less than 0.1% of feed-forward network neurons, that reliably predict when

Evian: Towards Explainable Visual Instruction-tuning Data Auditing

Model ReleasesDGX agent

arXiv:2604.20544v1 Announce Type: cross Abstract: The efficacy of Large Vision-Language Models (LVLMs) is critically dependent on the quality of their training data, requiring a precise balance betwee

Exploring Spatial Intelligence from a Generative Perspective

Model ReleasesDGX agent

arXiv:2604.20570v1 Announce Type: new Abstract: Spatial intelligence is essential for multimodal large language models, yet current benchmarks largely assess it only from an understanding perspective.

Generative Augmentation of Imbalanced Flight Records for Flight Diversion Prediction: A Multi-objective Optimisation Framework

SafetyDGX agent

arXiv:2604.20288v1 Announce Type: new Abstract: Flight diversions are rare but high-impact events in aviation, making their reliable prediction vital for both safety and operational efficiency. Howeve

LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning

Model ReleasesDGX agent

arXiv:2308.03303v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) is crucial for improving their performance on downstream tasks, but full-parameter fine-tuning (Full-FT) is

Self-Awareness before Action: Mitigating Logical Inertia via Proactive Cognitive Awareness

Model ReleasesDGX agent

arXiv:2604.20413v1 Announce Type: new Abstract: Large language models perform well on many reasoning tasks, yet they often lack awareness of whether their current knowledge or reasoning state is compl

Understanding the Staged Dynamics of Transformers in Learning Latent Structure

Model ReleasesDGX agent

arXiv:2511.19328v2 Announce Type: replace Abstract: Language modeling has shown us that transformers can discover latent structure from context, but the dynamics of how they acquire different componen

Wan-Image: Pushing the Boundaries of Generative Visual Intelligence

ApplicationsDGX agent

arXiv:2604.19858v1 Announce Type: new Abstract: We present Wan-Image, a unified visual generation system explicitly engineered to paradigm-shift image generation models from casual synthesizers into p

X-PCR: A Benchmark for Cross-modality Progressive Clinical Reasoning in Ophthalmic Diagnosis

Model ReleasesDGX agent

arXiv:2604.20350v1 Announce Type: new Abstract: Despite significant progress in Multi-modal Large Language Models (MLLMs), their clinical reasoning capacity for multi-modal diagnosis remains largely u

22 Apr 2026

Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications

SafetyDGX agent

arXiv:2604.19281v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) to support patients in addressing medical questions is becoming increasingly prevalent. However, most of the m

Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs

Model ReleasesDGX agent

arXiv:2604.18587v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated significant potential in formal theorem proving, yet state-of-the-art performance often necessitates pr

Council Mode: Mitigating Hallucination and Bias in LLMs via Multi-Agent Consensus

Model ReleasesDGX agent

arXiv:2604.02923v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly those employing Mixture-of-Experts (MoE) architectures, have achieved remarkable capabilities acros

Distillation Traps and Guards: A Calibration Knob for LLM Distillability

Local AiDGX agent

arXiv:2604.18963v1 Announce Type: cross Abstract: Knowledge distillation (KD) transfers capabilities from large language models (LLMs) to smaller students, yet it can fail unpredictably and also under

Gemma 4 VLA Demo on Jetson Orin Nano Super

Model ReleasesDGX agent

This article demonstrates running Gemma 4, Google's open-weight language model, on NVIDIA's Jetson Orin Nano Super edge computing device. It likely covers the model's capabilities, performance metrics

HALO: Hybrid Auto-encoded Locomotion with Learned Latent Dynamics, Poincare Maps, and Regions of Attraction

SafetyDGX agent

arXiv:2604.18887v1 Announce Type: new Abstract: Reduced-order models are powerful for analyzing and controlling high-dimensional dynamical systems. Yet constructing these models for complex hybrid sys

IMPACT: Importance-Aware Activation Space Reconstruction

ApplicationsDGX agent

arXiv:2507.03828v4 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse domains but remain difficult to deploy in resource-constrained environments d

Model-Agnostic Meta Learning for Class Imbalance Adaptation

ResearchDGX agent

arXiv:2604.18759v1 Announce Type: new Abstract: Class imbalance is a widespread challenge in NLP tasks, significantly hindering robust performance across diverse domains and applications. We introduce

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

SafetyDGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

Real-Time Streamable Generative Speech Restoration with Flow Matching

ResearchDGX agent

arXiv:2512.19442v3 Announce Type: replace-cross Abstract: Diffusion-based generative models have greatly impacted the speech processing field in recent years, exhibiting high speech naturalness and sp

SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension

Model ReleasesDGX agent

arXiv:2508.01959v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) over long documents typically involves splitting the text into smaller chunks, which serve as the basic units f

Towards Understanding the Robustness of Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2604.18756v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to optimization-based jailbreak attacks that exploit internal gradient structure. While Sparse Autoenco

UAF: A Unified Audio Front-end LLM for Full-Duplex Speech Interaction

ApplicationsDGX agent

arXiv:2604.19221v1 Announce Type: new Abstract: Full-duplex speech interaction, as the most natural and intuitive mode of human communication, is driving artificial intelligence toward more human-like

When Safety Fails Before the Answer: Benchmarking Harmful Behavior Detection in Reasoning Chains

Model ReleasesDGX agent

arXiv:2604.19001v1 Announce Type: new Abstract: Large reasoning models (LRMs) produce complex, multi-step reasoning traces, yet safety evaluation remains focused on final outputs, overlooking how harm

Who Shapes Brazil's Vaccine Debate? Semi-Supervised Modeling of Stance and Polarization in YouTube's Media Ecosystem

ResearchDGX agent

arXiv:2604.18586v1 Announce Type: cross Abstract: Vaccination remains a cornerstone of global public health, yet the COVID-19 pandemic exposed how online misinformation, political polarization, and de

21 Apr 2026

A Mechanism Study of Delayed Loss Spikes in Batch-Normalized Linear Models

ResearchDGX agent

arXiv:2604.16809v1 Announce Type: cross Abstract: Delayed loss spikes have been reported in neural-network training, but existing theory mainly explains earlier non-monotone behavior caused by overly

A Model and Estimation of the Bitcoin Transaction Fee

ResearchDGX agent

arXiv:2604.17183v1 Announce Type: cross Abstract: Bitcoin transaction fees will become more important as the block subsidy declines, but fee formation is hard to study with blockchain data alone becau

Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL

Model ReleasesDGX agent

arXiv:2604.17073v1 Announce Type: new Abstract: Reinforcement fine-tuning improves the reasoning ability of large language models, but it can also encourage them to answer unanswerable queries by gues

Adaptive Forensic Feature Refinement via Intrinsic Importance Perception

Model ReleasesDGX agent

arXiv:2604.16879v1 Announce Type: new Abstract: With the rapid development of generative models and multimodal content editing technologies, the key challenge faced by synthetic image detection (SID)

Are We Using the Right Benchmark: An Evaluation Framework for Visual Token Compression Methods

Model ReleasesDGX agent

arXiv:2510.07143v3 Announce Type: replace Abstract: Recent efforts to accelerate inference in Multimodal Large Language Models (MLLMs) have largely focused on visual token compression. The effectivene

Beyond Reproduction: A Paired-Task Framework for Assessing LLM Comprehension and Creativity in Literary Translation

Model ReleasesDGX agent

arXiv:2604.18169v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for creative tasks such as literary translation. Yet translational creativity remains underexplored a

Cat-DPO: Category-Adaptive Safety Alignment

SafetyDGX agent

arXiv:2604.17299v1 Announce Type: new Abstract: Aligning large language models with human preferences must balance two competing goals: responding helpfully to legitimate requests and reliably refusin

CoLLM: A Unified Framework for Co-execution of LLMs Federated Fine-tuning and Inference

Model ReleasesDGX agent

arXiv:2604.16400v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly adopted in edge intelligence to power domain-specific applications and personalized services, the qua

Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining

Model ReleasesDGX agent

arXiv:2604.16391v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have shown great potential in building generalist robots, but still face a dilemma-misalignment of 2D image foreca

eCP: Equivariant Conformal Prediction with pre-trained models

ResearchDGX agent

arXiv:2602.03986v2 Announce Type: replace Abstract: Conformal prediction, a post-hoc, distribution-free, finite-sample method of uncertainty quantification that offers formal coverage guarantees under

Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs

Model ReleasesDGX agent

arXiv:2510.11288v4 Announce Type: replace Abstract: Recent work has shown that narrow finetuning can produce broadly misaligned LLMs, a phenomenon termed emergent misalignment (EM). While concerning,

Forecasting Ionospheric Irregularities on GNSS Lines of Sight Using Dynamic Graphs with Ephemeris Conditioning

Model ReleasesDGX agent

arXiv:2604.18379v1 Announce Type: new Abstract: Most data-driven ionospheric forecasting models operate on gridded products, which do not preserve the time-varying sampling structure of satellite-base

FUSE: Ensembling Verifiers with Zero Labeled Data

ApplicationsDGX agent

arXiv:2604.18547v1 Announce Type: cross Abstract: Verification of model outputs is rapidly emerging as a key primitive for both training and real-world deployment of large language models (LLMs). In p

HiP-LoRA: Budgeted Spectral Plasticity for Robust Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2604.17751v1 Announce Type: cross Abstract: Adapting foundation models under resource budgets relies heavily on Parameter-Efficient Fine-Tuning (PEFT), with LoRA being a standard modular solutio

Kimi Kimi 。 Kimi Kimi Kimi Kimi Kimi Kimi ollama run kimi-k2.6:cloud

Local AiDGX agent

This appears to be a social media post from Ollama's X account regarding a model run command for 'kimi-k2.6:cloud,' likely announcing or demonstrating how to execute this specific AI model variant usi

Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes

ApplicationsDGX agent

arXiv:2604.18381v1 Announce Type: cross Abstract: Fine-tuning Large Language Models (LLMs) typically relies on large quantities of high-quality annotated data, or questions with well-defined ground tr

LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging

Model ReleasesDGX agent

arXiv:2511.07129v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has emerged as a parameter-efficient approach for fine-tuning large language models. However, conventional LoRA adapters

Medical thinking with multiple images

Model ReleasesDGX agent

arXiv:2604.16506v1 Announce Type: cross Abstract: Large language models perform well on many medical QA benchmarks, but real clinical reasoning often requires integrating evidence across multiple imag

Missing Pattern Tree based Decision Grouping and Ensemble for Enhancing Pair Utilization in Deep Incomplete Multi-View Clustering

Model ReleasesDGX agent

arXiv:2512.21510v2 Announce Type: replace-cross Abstract: Real-world multi-view data often exhibit highly inconsistent missing patterns, posing significant challenges for incomplete multi-view cluster

On the Importance and Evaluation of Narrativity in Natural Language AI Explanations

Model ReleasesDGX agent

arXiv:2604.18311v1 Announce Type: new Abstract: Explainable AI (XAI) aims to make the behaviour of machine learning models interpretable, yet many explanation methods remain difficult to understand. T

← Previous
1…265266267268269…1034
Next →