AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlog
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
Applications

Trustworthy deep domain adaptation for wearable photoplethysmography signal analysis with decision-theoretic uncertainty quantification

DGX agent

arXiv:2604.17480v1 Announce Type: new Abstract: In principle, deep generative models can be used to perform domain adaptation; i.e. align the input feature representations of test data with that of a

applicationsarxiv-cs-lg
21 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark

DGX agent

arXiv:2510.13759v3 Announce Type: replace Abstract: Unified multimodal models aim to jointly enable visual understanding and generation, yet current benchmarks rarely examine their true integration. E

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

1S-DAug: One-Shot Data Augmentation for Robust Few-Shot Generalization

DGX agent

arXiv:2602.00114v4 Announce Type: replace-cross Abstract: Few-shot learning (FSL) challenges model generalization to novel classes based on just a few shots of labeled examples, a testbed where tradit

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue

DGX agent

arXiv:2604.05552v2 Announce Type: replace-cross Abstract: Large Language Models demonstrate outstanding performance in many language tasks but still face fundamental challenges in managing the non-lin

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Facial-Expression-Aware Prompting for Empathetic LLM Tutoring

DGX agent

arXiv:2604.15336v1 Announce Type: cross Abstract: Large language models (LLMs) enable increasingly capable tutoring-style conversational agents, yet effective tutoring requires sensitivity to learners

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

From Benchmarking to Reasoning: A Dual-Aspect, Large-Scale Evaluation of LLMs on Vietnamese Legal Text

DGX agent

arXiv:2604.16270v1 Announce Type: cross Abstract: The complexity of Vietnam's legal texts presents a significant barrier to public access to justice. While Large Language Models offer a promising solu

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

FSPO: Few-Shot Optimization of Synthetic Preferences Personalizes to Real Users

DGX agent

arXiv:2502.19312v2 Announce Type: replace-cross Abstract: Effective personalization of LLMs is critical for a broad range of user-interfacing applications such as virtual assistants and content curati

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

Histogram-based Parameter-efficient Tuning for Passive and Active Sonar Classification

DGX agent

arXiv:2504.15214v3 Announce Type: replace Abstract: Parameter-efficient transfer learning (PETL) methods adapt large artificial neural networks to downstream tasks without fine-tuning the entire model

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge Distillation

DGX agent

arXiv:2505.13792v2 Announce Type: replace-cross Abstract: Recent advances in reasoning-focused Large Language Models (LLMs) have introduced Chain-of-Thought (CoT) traces - intermediate reasoning steps

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Kimi K2.6 now in OpenCode — Go included

DGX agent

Kimi K2.6, an AI model developed by Moonshot, has been released on OpenCode with support for the Go programming language. This release expands the model's capabilities to handle Go code generation, an

model-releaseskimi-moonshot--x
20 Apr 2026
Model Releases

LLMs Corrupt Your Documents When You Delegate

DGX agent

arXiv:2604.15597v1 Announce Type: new Abstract: Large Language Models (LLMs) are poised to disrupt knowledge work, with the emergence of delegated work as a new interaction paradigm (e.g., vibe coding

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Making Image Editing Easier via Adaptive Task Reformulation with Agentic Executions

DGX agent

arXiv:2604.15917v1 Announce Type: new Abstract: Instruction guided image editing has advanced substantially with recent generative models, yet it still fails to produce reliable results across many se

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Qwen3.5-Omni Technical Report

DGX agent

arXiv:2604.15804v1 Announce Type: new Abstract: In this work, we present Qwen3.5-Omni, the latest advancement in the Qwen-Omni model family. Representing a significant evolution over its predecessor,

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Self-Distillation as a Performance Recovery Mechanism for LLMs: Counteracting Compression and Catastrophic Forgetting

DGX agent

arXiv:2604.15794v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable success, underpinning diverse AI applications. However, they often suffer from performance degra

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

Stochasticity in Tokenisation Improves Robustness

DGX agent

arXiv:2604.16037v1 Announce Type: new Abstract: The widespread adoption of large language models (LLMs) has increased concerns about their robustness. Vulnerabilities in perturbations of tokenisation

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Time to really give Kimi K2.6 a go. Thank you @ollama! Love the ollama cloud setup!

DGX agent

Time to really give Kimi K2.6 a go. Thank you @ollama! Love the ollama cloud setup! Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch

model-releasesollama--x
20 Apr 2026
Model Releases

Transformer Neural Processes - Kernel Regression

DGX agent

arXiv:2411.12502v4 Announce Type: replace-cross Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic p

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

UniEditBench: A Unified and Cost-Effective Benchmark for Image and Video Editing via Distilled MLLMs

DGX agent

arXiv:2604.15871v1 Announce Type: cross Abstract: The evaluation of visual editing models remains fragmented across methods and modalities. Existing benchmarks are often tailored to specific paradigms

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

What Makes LLMs Effective Sequential Recommenders? A Study on Preference Intensity and Temporal Context

DGX agent

arXiv:2506.02261v3 Announce Type: replace-cross Abstract: What enables large language models (LLMs) to effectively model user preferences in sequential recommendation? Our investigation reveals that e

safetyarxiv-cs-lg
20 Apr 2026
Model Releases

Why Fine-Tuning Encourages Hallucinations and How to Fix It

DGX agent

arXiv:2604.15574v1 Announce Type: cross Abstract: Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information t

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

You ever go on Huggingface and see: - GGUF - Unsloth - Llama.cpp - Dynamic GGUF - Q_4_M / IQ_4XL etc. Here's what's going on under the hood.…

DGX agent

This post explains the technical details behind common terms and tools encountered on Hugging Face for running large language models locally, including quantization formats (GGUF, Q_4_M, IQ_4XL), opti

model-releasesclem-delangue--x
19 Apr 2026
Model Releases

Benchmarking Linguistic Adaptation in Comparable-Sized LLMs: A Study of Llama-3.1-8B, Mistral-7B-v0.1, and Qwen3-8B on Romanized Nepali

DGX agent

arXiv:2604.14171v1 Announce Type: new Abstract: Romanized Nepali, the Nepali language written in the Latin alphabet, is the dominant medium for informal digital communication in Nepal, yet it remains

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Between a Rock and a Hard Place: The Tension Between Ethical Reasoning and Safety Alignment in LLMs

DGX agent

arXiv:2509.05367v4 Announce Type: replace-cross Abstract: Large Language Model safety alignment predominantly operates on a binary assumption that requests are either safe or unsafe. This classificati

safetyarxiv-cs-ai
17 Apr 2026
Research

Beyond Translation: Evaluating Mathematical Reasoning Capabilities of LLMs in Sinhala and Tamil

DGX agent

arXiv:2602.14517v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong results in mathematical reasoning, and are increasingly deployed as tutoring and learning support

researcharxiv-cs-cl
17 Apr 2026
Safety

ConfLayers: Adaptive Confidence-based Layer Skipping for Self-Speculative Decoding

DGX agent

arXiv:2604.14612v1 Announce Type: cross Abstract: Self-speculative decoding is an inference technique for large language models designed to speed up generation without sacrificing output quality. It c

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

FoodSense: A Multisensory Food Dataset and Benchmark for Predicting Taste, Smell, Texture, and Sound from Images

DGX agent

arXiv:2604.14388v1 Announce Type: new Abstract: Humans routinely infer taste, smell, texture, and even sound from food images a phenomenon well studied in cognitive science. However, prior vision lang

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

Formalizing the Safety, Security, and Functional Properties of Agentic AI Systems

DGX agent

arXiv:2510.14133v2 Announce Type: replace Abstract: Agentic AI systems, which leverage multiple autonomous agents and large language models (LLMs), are increasingly used to address complex, multi-step

safetyarxiv-cs-ai
17 Apr 2026
Research

From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning

DGX agent

arXiv:2604.15244v1 Announce Type: new Abstract: Speculative decoding (SD) accelerates large language model inference by allowing a lightweight draft model to propose outputs that a stronger target mod

researcharxiv-cs-cl
17 Apr 2026
Safety

Integrating Object Detection, LiDAR-Enhanced Depth Estimation, and Segmentation Models for Railway Environments

DGX agent

arXiv:2604.14781v1 Announce Type: new Abstract: Obstacle detection in railway environments is crucial for ensuring safety. However, very few studies address the problem using a complete, modular, and

safetyarxiv-cs-cv
17 Apr 2026
Model Releases

Magnitude Is All You Need? Rethinking Phase in Quantum Encoding of Complex SAR Data

DGX agent

arXiv:2604.14229v1 Announce Type: cross Abstract: Synthetic Aperture Radar (SAR) data is inherently complex-valued, while quantum machine learning (QML) models naturally operate in complex Hilbert spa

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

MARCA: A Checklist-Based Benchmark for Multilingual Web Search

DGX agent

arXiv:2604.14448v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as sources of information, yet their reliability depends on the ability to search the web, select rel

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Mechanistic Decoding of Cognitive Constructs in LLMs

DGX agent

arXiv:2604.14593v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate increasingly sophisticated affective capabilities, the internal mechanisms by which they process complex

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

MIND: AI Co-Scientist for Material Research

DGX agent

arXiv:2604.13699v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled agentic AI systems for scientific discovery, but most approaches remain limited to textbased reasoning witho

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

Physically-Induced Atmospheric Adversarial Perturbations: Enhancing Transferability and Robustness in Remote Sensing Image Classification

DGX agent

arXiv:2604.14643v1 Announce Type: new Abstract: Adversarial attacks pose a severe threat to the reliability of deep learning models in remote sensing (RS) image classification. Most existing methods r

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Shuffle the Context: RoPE-Perturbed Self-Distillation for Long-Context Adaptation

DGX agent

arXiv:2604.14339v1 Announce Type: new Abstract: Large language models (LLMs) increasingly operate in settings that require reliable long-context understanding, such as retrieval-augmented generation a

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

SOLIS: Physics-Informed Learning of Interpretable Neural Surrogates for Nonlinear Systems

DGX agent

arXiv:2604.14879v1 Announce Type: new Abstract: Nonlinear system identification must balance physical interpretability with model flexibility. Classical methods yield structured, control-relevant mode

model-releasesarxiv-cs-lg
17 Apr 2026
Applications

TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation

DGX agent

arXiv:2604.14580v1 Announce Type: new Abstract: Existing audio-driven video digital human generation models rely on multi-step denoising, resulting in substantial computational overhead that severely

applicationsarxiv-cs-cv
17 Apr 2026
Model Releases

When Flat Minima Fail: Characterizing INT4 Quantization Collapse After FP32 Convergence

DGX agent

arXiv:2604.15167v1 Announce Type: new Abstract: Post-training quantization (PTQ) assumes that a well-converged model is a quantization-ready model. We show this assumption fails in a structured, measu

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

Enhancing Confidence Estimation in Telco LLMs via Twin-Pass CoT-Ensembling

DGX agent

arXiv:2604.13271v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly applied to complex telecommunications tasks, including 3GPP specification analysis and O-RAN network troub

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Free Geometry: Refining 3D Reconstruction from Longer Versions of Itself

DGX agent

arXiv:2604.14048v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models are efficient but rigid: once trained, they perform inference in a zero-shot manner and cannot adapt to the test s

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

From Weights to Activations: Is Steering the Next Frontier of Adaptation?

DGX agent

arXiv:2604.14090v1 Announce Type: new Abstract: Post-training adaptation of language models is commonly achieved through parameter updates or input-based methods such as fine-tuning, parameter-efficie

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Go from blank slate to analysis with BigQuery Studio notebook gallery templates

DGX agent

For many data professionals, the most daunting part of a new project isn't the complexity of the data or the sophistication of the model, it’s the 'blank slate.' Staring at an empty notebook while cre

model-releasesgoogle-cloud-ai
16 Apr 2026
Model Releases

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Re…

DGX agent

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Refusal rate: 98.8% down to 2.1% (!!) → Compliance: 1.2% up to

model-releasesclem-delangue--x
16 Apr 2026
Research

Identifiability of Potentially Degenerate Gaussian Mixture Models With Piecewise Affine Mixing

DGX agent

arXiv:2604.13218v1 Announce Type: cross Abstract: Causal representation learning (CRL) aims to identify the underlying latent variables from high-dimensional observations, even when variables are depe

researcharxiv-cs-lg
16 Apr 2026
Model Releases

Introducing GPT-Rosalind for life sciences research

DGX agent

GPT-Rosalind is an AI model developed by OpenAI specifically designed to assist with life sciences research tasks. The model is trained to help researchers with applications such as analyzing biologic

model-releasesopenai
16 Apr 2026
Applications

Mathematical Reasoning Enhanced LLM for Formula Derivation: A Case Study on Fiber NLI Modellin

DGX agent

arXiv:2604.13062v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have demonstrated strong capabilities in code generation and text synthesis, yet their potential for sym

applicationsarxiv-cs-cl
16 Apr 2026
Model Releases

Mitigating Catastrophic Forgetting in Target Language Adaptation of LLMs via Source-Shielded Updates

DGX agent

arXiv:2512.04844v2 Announce Type: replace Abstract: Expanding the linguistic diversity of instruct large language models (LLMs) is crucial for global accessibility but is often hindered by the relianc

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Qwen 3.6 is here, and open-source! Run it locally with improved agentic coding capabilities. Try it with Claude Code: ollama launch claude -…

DGX agent

Qwen 3.6 is here, and open-source! Run it locally with improved agentic coding capabilities. Try it with Claude Code: ollama launch claude --model qwen3.6 Try it with OpenClaw: ollama launch openclaw

model-releasesollama--x
16 Apr 2026
← Previous
1…372373374375376…1324
Next →