AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlog
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,509 results
Safety

Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications

DGX agent

arXiv:2604.19281v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) to support patients in addressing medical questions is becoming increasingly prevalent. However, most of the m

safetyarxiv-cs-ai
22 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs

DGX agent

arXiv:2604.18587v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated significant potential in formal theorem proving, yet state-of-the-art performance often necessitates pr

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Council Mode: Mitigating Hallucination and Bias in LLMs via Multi-Agent Consensus

DGX agent

arXiv:2604.02923v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly those employing Mixture-of-Experts (MoE) architectures, have achieved remarkable capabilities acros

model-releasesarxiv-cs-ai
22 Apr 2026
Local Ai

Distillation Traps and Guards: A Calibration Knob for LLM Distillability

DGX agent

arXiv:2604.18963v1 Announce Type: cross Abstract: Knowledge distillation (KD) transfers capabilities from large language models (LLMs) to smaller students, yet it can fail unpredictably and also under

local-aiarxiv-cs-ai
22 Apr 2026
Model Releases

Gemma 4 VLA Demo on Jetson Orin Nano Super

DGX agent

This article demonstrates running Gemma 4, Google's open-weight language model, on NVIDIA's Jetson Orin Nano Super edge computing device. It likely covers the model's capabilities, performance metrics

model-releaseshugging-face
22 Apr 2026
Safety

HALO: Hybrid Auto-encoded Locomotion with Learned Latent Dynamics, Poincare Maps, and Regions of Attraction

DGX agent

arXiv:2604.18887v1 Announce Type: new Abstract: Reduced-order models are powerful for analyzing and controlling high-dimensional dynamical systems. Yet constructing these models for complex hybrid sys

safetyarxiv-cs-ro
22 Apr 2026
Applications

IMPACT: Importance-Aware Activation Space Reconstruction

DGX agent

arXiv:2507.03828v4 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse domains but remain difficult to deploy in resource-constrained environments d

applicationsarxiv-cs-lg
22 Apr 2026
Research

Model-Agnostic Meta Learning for Class Imbalance Adaptation

DGX agent

arXiv:2604.18759v1 Announce Type: new Abstract: Class imbalance is a widespread challenge in NLP tasks, significantly hindering robust performance across diverse domains and applications. We introduce

researcharxiv-cs-cl
22 Apr 2026
Safety

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

DGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

safetyarxiv-cs-cl
22 Apr 2026
Research

Real-Time Streamable Generative Speech Restoration with Flow Matching

DGX agent

arXiv:2512.19442v3 Announce Type: replace-cross Abstract: Diffusion-based generative models have greatly impacted the speech processing field in recent years, exhibiting high speech naturalness and sp

researcharxiv-cs-lg
22 Apr 2026
Model Releases

SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension

DGX agent

arXiv:2508.01959v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) over long documents typically involves splitting the text into smaller chunks, which serve as the basic units f

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Towards Understanding the Robustness of Sparse Autoencoders

DGX agent

arXiv:2604.18756v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to optimization-based jailbreak attacks that exploit internal gradient structure. While Sparse Autoenco

model-releasesarxiv-cs-ai
22 Apr 2026
Applications

UAF: A Unified Audio Front-end LLM for Full-Duplex Speech Interaction

DGX agent

arXiv:2604.19221v1 Announce Type: new Abstract: Full-duplex speech interaction, as the most natural and intuitive mode of human communication, is driving artificial intelligence toward more human-like

applicationsarxiv-cs-ai
22 Apr 2026
Model Releases

When Safety Fails Before the Answer: Benchmarking Harmful Behavior Detection in Reasoning Chains

DGX agent

arXiv:2604.19001v1 Announce Type: new Abstract: Large reasoning models (LRMs) produce complex, multi-step reasoning traces, yet safety evaluation remains focused on final outputs, overlooking how harm

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Who Shapes Brazil's Vaccine Debate? Semi-Supervised Modeling of Stance and Polarization in YouTube's Media Ecosystem

DGX agent

arXiv:2604.18586v1 Announce Type: cross Abstract: Vaccination remains a cornerstone of global public health, yet the COVID-19 pandemic exposed how online misinformation, political polarization, and de

researcharxiv-cs-ai
22 Apr 2026
Research

A Mechanism Study of Delayed Loss Spikes in Batch-Normalized Linear Models

DGX agent

arXiv:2604.16809v1 Announce Type: cross Abstract: Delayed loss spikes have been reported in neural-network training, but existing theory mainly explains earlier non-monotone behavior caused by overly

researcharxiv-cs-lg
21 Apr 2026
Research

A Model and Estimation of the Bitcoin Transaction Fee

DGX agent

arXiv:2604.17183v1 Announce Type: cross Abstract: Bitcoin transaction fees will become more important as the block subsidy declines, but fee formation is hard to study with blockchain data alone becau

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL

DGX agent

arXiv:2604.17073v1 Announce Type: new Abstract: Reinforcement fine-tuning improves the reasoning ability of large language models, but it can also encourage them to answer unanswerable queries by gues

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Adaptive Forensic Feature Refinement via Intrinsic Importance Perception

DGX agent

arXiv:2604.16879v1 Announce Type: new Abstract: With the rapid development of generative models and multimodal content editing technologies, the key challenge faced by synthetic image detection (SID)

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Are We Using the Right Benchmark: An Evaluation Framework for Visual Token Compression Methods

DGX agent

arXiv:2510.07143v3 Announce Type: replace Abstract: Recent efforts to accelerate inference in Multimodal Large Language Models (MLLMs) have largely focused on visual token compression. The effectivene

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Beyond Reproduction: A Paired-Task Framework for Assessing LLM Comprehension and Creativity in Literary Translation

DGX agent

arXiv:2604.18169v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for creative tasks such as literary translation. Yet translational creativity remains underexplored a

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Cat-DPO: Category-Adaptive Safety Alignment

DGX agent

arXiv:2604.17299v1 Announce Type: new Abstract: Aligning large language models with human preferences must balance two competing goals: responding helpfully to legitimate requests and reliably refusin

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

CoLLM: A Unified Framework for Co-execution of LLMs Federated Fine-tuning and Inference

DGX agent

arXiv:2604.16400v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly adopted in edge intelligence to power domain-specific applications and personalized services, the qua

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining

DGX agent

arXiv:2604.16391v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have shown great potential in building generalist robots, but still face a dilemma-misalignment of 2D image foreca

model-releasesarxiv-cs-cv
21 Apr 2026
Research

eCP: Equivariant Conformal Prediction with pre-trained models

DGX agent

arXiv:2602.03986v2 Announce Type: replace Abstract: Conformal prediction, a post-hoc, distribution-free, finite-sample method of uncertainty quantification that offers formal coverage guarantees under

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs

DGX agent

arXiv:2510.11288v4 Announce Type: replace Abstract: Recent work has shown that narrow finetuning can produce broadly misaligned LLMs, a phenomenon termed emergent misalignment (EM). While concerning,

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Forecasting Ionospheric Irregularities on GNSS Lines of Sight Using Dynamic Graphs with Ephemeris Conditioning

DGX agent

arXiv:2604.18379v1 Announce Type: new Abstract: Most data-driven ionospheric forecasting models operate on gridded products, which do not preserve the time-varying sampling structure of satellite-base

model-releasesarxiv-cs-lg
21 Apr 2026
Applications

FUSE: Ensembling Verifiers with Zero Labeled Data

DGX agent

arXiv:2604.18547v1 Announce Type: cross Abstract: Verification of model outputs is rapidly emerging as a key primitive for both training and real-world deployment of large language models (LLMs). In p

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

HiP-LoRA: Budgeted Spectral Plasticity for Robust Low-Rank Adaptation

DGX agent

arXiv:2604.17751v1 Announce Type: cross Abstract: Adapting foundation models under resource budgets relies heavily on Parameter-Efficient Fine-Tuning (PEFT), with LoRA being a standard modular solutio

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Kimi Kimi 。 Kimi Kimi Kimi Kimi Kimi Kimi ollama run kimi-k2.6:cloud

DGX agent

This appears to be a social media post from Ollama's X account regarding a model run command for 'kimi-k2.6:cloud,' likely announcing or demonstrating how to execute this specific AI model variant usi

local-aiollama--x
21 Apr 2026
Applications

Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes

DGX agent

arXiv:2604.18381v1 Announce Type: cross Abstract: Fine-tuning Large Language Models (LLMs) typically relies on large quantities of high-quality annotated data, or questions with well-defined ground tr

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging

DGX agent

arXiv:2511.07129v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has emerged as a parameter-efficient approach for fine-tuning large language models. However, conventional LoRA adapters

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Medical thinking with multiple images

DGX agent

arXiv:2604.16506v1 Announce Type: cross Abstract: Large language models perform well on many medical QA benchmarks, but real clinical reasoning often requires integrating evidence across multiple imag

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Missing Pattern Tree based Decision Grouping and Ensemble for Enhancing Pair Utilization in Deep Incomplete Multi-View Clustering

DGX agent

arXiv:2512.21510v2 Announce Type: replace-cross Abstract: Real-world multi-view data often exhibit highly inconsistent missing patterns, posing significant challenges for incomplete multi-view cluster

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

On the Importance and Evaluation of Narrativity in Natural Language AI Explanations

DGX agent

arXiv:2604.18311v1 Announce Type: new Abstract: Explainable AI (XAI) aims to make the behaviour of machine learning models interpretable, yet many explanation methods remain difficult to understand. T

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

PAC-Bayes Bounds for Gibbs Posteriors via Singular Learning Theory

DGX agent

arXiv:2604.17219v1 Announce Type: cross Abstract: We derive explicit non-asymptotic PAC-Bayes generalization bounds for Gibbs posteriors, that is, data-dependent distributions over model parameters ob

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

PBSBench: A Multi-Level Vision-Language Framework and Benchmark for Hematopathology Whole Slide Image Interpretation

DGX agent

arXiv:2604.17570v1 Announce Type: new Abstract: Peripheral Blood Smear (PBS) is a critical microscopic examination in hematopathology that yields whole-slide imaging (WSI). Unlike solid tissue patholo

model-releasesarxiv-cs-cv
21 Apr 2026
Research

PCM-NeRF: Probabilistic Camera Modeling for Neural Radiance Fields under Pose Uncertainty

DGX agent

arXiv:2604.17831v1 Announce Type: new Abstract: Neural surface reconstruction methods typically treat camera poses as fixed values, assuming perfect accuracy from Structure-from-Motion (SfM) systems.

researcharxiv-cs-cv
21 Apr 2026
Model Releases

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning

DGX agent

arXiv:2604.17800v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have gained much attention from the research community thanks to their strength in translating multimodal observat

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

RosettaSearch: Multi-Objective Inference-Time Search for Protein Sequence Design

DGX agent

arXiv:2604.17175v1 Announce Type: new Abstract: We introduce RosettaSearch, an inference-time multi-objective optimization approach for protein sequence optimization. We use large language models (LLM

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

SetFlow: Generating Structured Sets of Representations for Multiple Instance Learning

DGX agent

arXiv:2604.16362v1 Announce Type: cross Abstract: Data scarcity and weak supervision continue to limit the performance of machine learning models in many real-world applications, such as mammography,

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Tool Learning Needs Nothing More Than a Free 8B Language Model

DGX agent

arXiv:2604.17739v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a prevalent paradigm for training tool calling agents, which typically requires online interactive environments

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Universally Empowering Zeroth-Order Optimization via Adaptive Layer-wise Sampling

DGX agent

arXiv:2604.18264v1 Announce Type: new Abstract: Zeroth-Order optimization presents a promising memory-efficient paradigm for fine-tuning Large Language Models by relying solely on forward passes. Howe

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Beyond Distribution Sharpening: The Importance of Task Rewards

DGX agent

arXiv:2604.16259v1 Announce Type: cross Abstract: Frontier models have demonstrated exceptional capabilities following the integration of task-reward-based reinforcement learning (RL) into their train

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Beyond MCQ: An Open-Ended Arabic Cultural QA Benchmark with Dialect Variants

DGX agent

arXiv:2510.24328v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used to answer everyday questions, yet their performance on culturally grounded and dialectal co

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Brain Score Tracks Shared Properties of Languages: Evidence from Many Natural Languages and Structured Sequences

DGX agent

arXiv:2604.15503v1 Announce Type: new Abstract: Recent breakthroughs in language models (LMs) using neural networks have raised the question: how similar are these models' processing to human language

researcharxiv-cs-cl
20 Apr 2026
Model Releases

Frequency-Aware Flow Matching for High-Quality Image Generation

DGX agent

arXiv:2604.15521v1 Announce Type: new Abstract: Flow matching models have emerged as a powerful framework for realistic image generation by learning to reverse a corruption process that progressively

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

Long-Term Memory for VLA-based Agents in Open-World Task Execution

DGX agent

arXiv:2604.15671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated significant potential for embodied decision-making; however, their application in complex chemical

safetyarxiv-cs-ro
20 Apr 2026
← Previous
1…335336337338339…1303
Next →