AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlog
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,172 results
Agents

Bi-Predictability: A Real-Time Signal for Monitoring LLM Interaction Integrity

DGX agent

arXiv:2604.13061v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in high-stakes autonomous and interactive workflows, where reliability demands continuous, multi-

agentsarxiv-cs-cl
16 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Breaking the Generator Barrier: Disentangled Representation for Generalizable AI-Text Detection

DGX agent

arXiv:2604.13692v1 Announce Type: new Abstract: As large language models (LLMs) generate text that increasingly resembles human writing, the subtle cues that distinguish AI-generated content from huma

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Claude Opus 4.7 is now available as an Agent Preview inside of Devin! Anthropic has clearly optimized Claude Opus 4.7 for long-horizon auton…

DGX agent

Claude Opus 4.7 is now available as an Agent Preview inside of Devin! Anthropic has clearly optimized Claude Opus 4.7 for long-horizon autonomy, unlocking a class of deep investigation work we couldn'

model-releasescognition-ai--x
16 Apr 2026
Model Releases

Codex for (almost) everything

DGX agent

OpenAI's Codex is a large language model trained on publicly available code from the internet that can understand and generate code in dozens of programming languages. It powers GitHub Copilot and can

model-releasesopenai
16 Apr 2026
Model Releases

Counterfactual Peptide Editing for Causal TCR--pMHC Binding Inference

DGX agent

arXiv:2604.13256v1 Announce Type: new Abstract: Neural models for TCR-pMHC binding prediction are susceptible to shortcut learning: they exploit spurious correlations in training data -- such as pepti

model-releasesarxiv-cs-lg
16 Apr 2026
Research

English is Not All You Need: Systematically Exploring the Role of Multilinguality in LLM Post-Training

DGX agent

arXiv:2604.13286v1 Announce Type: new Abstract: Despite the widespread multilingual deployment of large language models, post-training pipelines remain predominantly English-centric, contributing to p

researcharxiv-cs-cl
16 Apr 2026
Safety

Estimating Continuous Treatment Effects with Two-Stage Kernel Ridge Regression

DGX agent

arXiv:2604.13410v1 Announce Type: cross Abstract: We study the problem of estimating the effect function for a continuous treatment, which maps each treatment value to a population-averaged outcome. A

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

Exposia: Teaching and Assessment of Academic Writing Skills for Research Project Proposals and Peer Feedback

DGX agent

arXiv:2601.06536v2 Announce Type: replace Abstract: We present Exposia, the first public dataset that connects writing and feedback in higher education, enabling research on educationally grounded com

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

FiLM-Nav: Efficient and Generalizable Navigation via VLM Fine-tuning

DGX agent

arXiv:2509.16445v2 Announce Type: replace Abstract: Enabling robotic assistants to navigate complex environments and locate objects described in free-form language is a critical capability for real-wo

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

FlexGuard: Continuous Risk Scoring for Strictness-Adaptive LLM Content Moderation

DGX agent

arXiv:2602.23636v3 Announce Type: replace Abstract: Ensuring the safety of LLM-generated content is essential for real-world deployment. Most existing guardrail models formulate moderation as a fixed

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Functional Emotions or Situational Contexts? A Discriminating Test from the Mythos Preview System Card

DGX agent

arXiv:2604.13466v1 Announce Type: cross Abstract: The Claude Mythos Preview system card deploys emotion vectors, sparse autoencoder (SAE) features, and activation verbalisers to study model internals

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

GLM-5.1 Tool Calling Issue Fix & Chat Template Update If you are running GLM-5.1 with vLLM/SGLang and using tool calling, please update your…

DGX agent

GLM-5.1 Tool Calling Issue Fix & Chat Template Update If you are running GLM-5.1 with vLLM/SGLang and using tool calling, please update your chat template. http://huggingface.co/zai-org/GLM-5.1/blob/m

model-releaseszhipu-ai--x
16 Apr 2026
Model Releases

Here's Qwen 3.6-35B-A3B v.s. Claude Opus 4.7 for 'Generate an SVG of a flamingo riding a unicycle', in case you thought Qwen might be cheati…

DGX agent

This post compares the performance of Qwen 3.6-35B-A3B and Claude Opus 4.7 models on a creative task of generating SVG code for a flamingo riding a unicycle, likely demonstrating differences in their

model-releasessimon-willison--x
16 Apr 2026
Model Releases

KMMMU: Evaluation of Massive Multi-discipline Multimodal Understanding in Korean Language and Context

DGX agent

arXiv:2604.13058v1 Announce Type: new Abstract: We introduce KMMMU, a native Korean benchmark for evaluating multimodal understanding in Korean cultural and institutional settings. KMMMU contains 3,46

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs

DGX agent

arXiv:2604.13226v1 Announce Type: new Abstract: Large Language Models (LLMs) rely heavily on Key-Value (KV) caching to minimize inference latency. However, standard KV caches are context-dependent: re

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Language steering in latent space to mitigate unintended code-switching

DGX agent

arXiv:2510.13849v3 Announce Type: replace Abstract: Multilingual Large Language Models (LLMs) often exhibit hallucinations such as unintended code-switching, reducing reliability in downstream tasks.

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

LM Performance:Qwen3.6-35B-A3B outperforms the dense 27B-param Qwen3.5-27B on several key coding benchmarks and dramatically surpasses its d…

DGX agent

LM Performance:Qwen3.6-35B-A3B outperforms the dense 27B-param Qwen3.5-27B on several key coding benchmarks and dramatically surpasses its direct predecessor Qwen3.5-35B-A3B, especially on agentic cod

agentsqwen--x
16 Apr 2026
Local Ai

LTX distilled 1.1 is the new king!

DGX agent

A Reddit thread from r/StableDiffusion discussing the release and community reception of LTX-Video's distilled 1.1 model, developed by Lightricks. LTX-Video is described as the first DiT-based video g

local-air-stablediffusion
16 Apr 2026
Local Ai

Need help setting up ollama.

DGX agent

A Reddit thread from the r/ollama community where a user seeks assistance with the initial setup and configuration of Ollama, a tool for running large language models locally. The discussion likely co

local-air-ollama
16 Apr 2026
Local Ai

Nucleus Image now supported in Ostris' AI-Toolkit.

DGX agent

Ostris' AI Toolkit is an all-in-one training suite for diffusion models , and Nucleus Image has been added to the list of supported models . The toolkit can be run as a GUI or CLI and is designed to b

local-air-stablediffusion
16 Apr 2026
Research

On an L^2 norm for stationary ARMA processes

DGX agent

arXiv:2408.10610v5 Announce Type: replace Abstract: We propose an L^2 norm for stationary Autoregressive Moving Average (ARMA) models. We look at ARMA models within the Hilbert space of the past with

researcharxiv-cs-lg
16 Apr 2026
Model Releases

Online learning with noisy side observations

DGX agent

arXiv:2604.13740v1 Announce Type: new Abstract: We propose a new partial-observability model for online learning problems where the learner, besides its own loss, also observes some noisy feedback abo

model-releasesarxiv-cs-lg
16 Apr 2026
Industry

OpenAI starts offering a biology-tuned LLM

DGX agent

OpenAI has launched GPT-Rosalind, a biology-tuned large language model . This specialized LLM is designed to enhance performance on biology-specific tasks and applications. The model represents OpenAI

industryars-technica
16 Apr 2026
Model Releases

Optimization with SpotOptim

DGX agent

arXiv:2604.13672v1 Announce Type: new Abstract: The `spotoptim` package implements surrogate-model-based optimization of expensive black-box functions in Python. Building on two decades of Sequential

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Parameter Importance is Not Static: Evolving Parameter Isolation for Supervised Fine-Tuning

DGX agent

arXiv:2604.14010v1 Announce Type: cross Abstract: Supervised Fine-Tuning (SFT) of large language models often suffers from task interference and catastrophic forgetting. Recent approaches alleviate th

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

PBE-UNet: A light weight Progressive Boundary-Enhanced U-Net with Scale-Aware Aggregation for Ultrasound Image Segmentation

DGX agent

arXiv:2604.13791v1 Announce Type: new Abstract: Accurate lesion segmentation in ultrasound images is essential for preventive screening and clinical diagnosis, yet remains challenging due to low contr

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

PersonaVLM: Long-Term Personalized Multimodal LLMs

DGX agent

arXiv:2604.13074v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) serve as daily assistants for millions. However, their ability to generate responses aligned with individual pr

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

RL-PLUS: Countering Capability Boundary Collapse of LLMs in Reinforcement Learning with Hybrid-policy Optimization

DGX agent

arXiv:2508.00222v5 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Reward (RLVR) has significantly advanced the complex reasoning abilities of Large Language Models (LLMs

safetyarxiv-cs-cl
16 Apr 2026
Model Releases

Seedance 2.0: Advancing Video Generation for World Complexity

DGX agent

arXiv:2604.14148v1 Announce Type: new Abstract: Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecesso

model-releasesarxiv-cs-cv
16 Apr 2026
Research

SiLVR: A Simple Language-based Video Reasoning Framework

DGX agent

arXiv:2505.24869v3 Announce Type: replace Abstract: Recent advances in test-time optimization have led to remarkable reasoning capabilities in Large Language Models (LLMs), enabling them to solve high

researcharxiv-cs-cv
16 Apr 2026
Model Releases

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs

DGX agent

arXiv:2604.13710v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) exhibit strong reasoning and world knowledge, yet adapting them for retrieval remains challenging. Existing app

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing!

DGX agent

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing! ⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agen

model-releasesclem-delangue--x
16 Apr 2026
Research

The Signal is in the Steps: Local Scoring for Reasoning Data Selection

DGX agent

arXiv:2510.03988v2 Announce Type: replace Abstract: Distilling long-form reasoning from teacher models into smaller students requires selecting which candidate solutions to train on. Recent work argue

researcharxiv-cs-lg
16 Apr 2026
Tutorials

Tokenizing Semantic Segmentation with Run Length Encoding

DGX agent

arXiv:2602.21627v3 Announce Type: replace Abstract: This paper presents a new unified approach to semantic segmentation in both images and videos by using language modeling to output the masks as sequ

tutorialsarxiv-cs-cv
16 Apr 2026
Model Releases

Two-Stage Regularization-Based Structured Pruning for LLMs

DGX agent

arXiv:2505.18232v3 Announce Type: replace-cross Abstract: The deployment of large language models (LLMs) is largely hindered by their large number of parameters. Structural pruning has emerged as a pr

model-releasesarxiv-cs-cl
16 Apr 2026
Local Ai

Unleashing Implicit Rewards: Prefix-Value Learning for Distribution-Level Optimization

DGX agent

arXiv:2604.13197v1 Announce Type: new Abstract: Process reward models (PRMs) provide fine-grained reward signals along the reasoning process, but training reliable PRMs often requires step annotations

local-aiarxiv-cs-cl
16 Apr 2026
Model Releases

What’s New in Microsoft Foundry Fine-Tuning | April 2026

DGX agent

April 2026 brings three major Reinforcement Fine-Tuning updates: Global Training for o4-mini with lower per-token rates across 12+ regions, new GPT-4.1 model graders for richer reward signals, and a c

model-releasesmicrosoft-foundry
16 Apr 2026
Model Releases

WorkRB: A Community-Driven Evaluation Framework for AI in the Work Domain

DGX agent

arXiv:2604.13055v1 Announce Type: new Abstract: Today's evolving labor markets rely increasingly on recommender systems for hiring, talent management, and workforce analytics, with natural language pr

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

A Sanity Check on Composed Image Retrieval

DGX agent

arXiv:2604.12904v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) aims to retrieve a target image based on a query composed of a reference image, and a relative caption that specifies the

model-releasesarxiv-cs-cv
15 Apr 2026
Research

AdaMCoT: Rethinking Cross-Lingual Factual Reasoning through Adaptive Multilingual Chain-of-Thought

DGX agent

arXiv:2501.16154v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown impressive multilingual capabilities through pretraining on diverse corpora. Although these models sho

researcharxiv-cs-ai
15 Apr 2026
Model Releases

AISafetyBenchExplorer: A Metric-Aware Catalogue of AI Safety Benchmarks Reveals Fragmented Measurement and Weak Benchmark Governance

DGX agent

arXiv:2604.12875v1 Announce Type: new Abstract: The rapid expansion of large language model (LLM) safety evaluation has produced a substantial benchmark ecosystem, but not a correspondingly coherent m

model-releasesarxiv-cs-ai
15 Apr 2026
Agents

CascadeDebate: Multi-Agent Deliberation for Cost-Aware LLM Cascades

DGX agent

arXiv:2604.12262v1 Announce Type: cross Abstract: Cascaded LLM systems coordinate models of varying sizes with human experts to balance accuracy, cost, and abstention under uncertainty. However, singl

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

CoD-Lite: Real-Time Diffusion-Based Generative Image Compression

DGX agent

arXiv:2604.12525v1 Announce Type: new Abstract: Recent advanced diffusion methods typically derive strong generative priors by scaling diffusion transformers. However, scaling fails to generalize when

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

CoDe-R: Refining Decompiler Output with LLMs via Rationale Guidance and Adaptive Inference

DGX agent

arXiv:2604.12913v1 Announce Type: cross Abstract: Binary decompilation is a critical reverse engineering task aimed at reconstructing high-level source code from stripped executables. Although Large L

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation

DGX agent

arXiv:2604.12268v1 Announce Type: cross Abstract: Large language models (LLMs) can generate code from natural language, but the extent to which they capture intended program behavior remains unclear.

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

FAST-DIPS: Adjoint-Free Analytic Steps and Hard-Constrained Likelihood Correction for Diffusion-Prior Inverse Problems

DGX agent

arXiv:2603.01591v2 Announce Type: replace-cross Abstract: Training-free diffusion priors enable inverse-problem solvers without retraining, but for nonlinear forward operators data consistency often r

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

FMASH: Advancing Traditional Chinese Medicine Formula Recommendation with Efficient Fusion of Multiscale Associations of Symptoms and Herbs

DGX agent

arXiv:2503.05167v3 Announce Type: replace Abstract: Traditional Chinese medicine (TCM) exhibits remarkable therapeutic efficacy in healthcare through patient-specific formulas. However, current AI-bas

local-aiarxiv-cs-lg
15 Apr 2026
Model Releases

From Imitation to Discrimination: Progressive Curriculum Learning for Robust Web Navigation

DGX agent

arXiv:2604.12666v1 Announce Type: cross Abstract: Text-based web agents offer computational efficiency for autonomous web navigation, yet developing robust agents remains challenging due to the noisy

model-releasesarxiv-cs-cl
15 Apr 2026
← Previous
1…477478479480481…1358
Next →