AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,082 results
13 Apr 2026

SAGE: A Service Agent Graph-guided Evaluation Benchmark

Model ReleasesDGX agent

arXiv:2604.09285v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has catalyzed automation in customer service, yet benchmarking their performance remains challenging. Ex

Skill-Conditioned Visual Geolocation for Vision-Language

Model ReleasesDGX agent

arXiv:2604.09025v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown a promising ability in image geolocation, but they still lack structured geographic reasoning and the capacit

So the concern over Mythos and cybersecurity seems warranted.

Model ReleasesDGX agent

So the concern over Mythos and cybersecurity seems warranted. We conducted cyber evaluations of Claude Mythos Preview and found that it is the first model to complete an AISI cyber range end-to-end. 🧵

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SUPERNOVA: Eliciting General Reasoning in LLMs with Reinforcement Learning on Natural Instructions

ResearchDGX agent

arXiv:2604.08477v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has significantly improved large language model (LLM) reasoning in formal domains such as mathem

Turning Anime into Real and testing Klein9b vs Qwen Edit 2511 (Workflow Included)

Model ReleasesDGX agent

This r/StableDiffusion post showcases a workflow for converting anime-style images into photorealistic outputs, directly comparing two AI image editing models: Klein 9B and Qwen Image Edit 2511. Qwen

UIPress: Bringing Optical Token Compression to UI-to-Code Generation

ResearchDGX agent

arXiv:2604.09442v1 Announce Type: new Abstract: UI-to-Code generation requires vision-language models (VLMs) to produce thousands of tokens of structured HTML/CSS from a single screenshot, making visu

You Can't Fight in Here! This is BBS!

TutorialsDGX agent

arXiv:2604.09501v1 Announce Type: new Abstract: Norm, the formal theoretical linguist, and Claudette, the computational language scientist, have a lovely time discussing whether modern language models

12 Apr 2026

Amazing article. If you own the harness you own your memories. Else you get locked into an API model which keeps memory behind APIs and you …

AgentsDGX agent

Harrison Chase, co-founder of LangChain, shared or engaged with commentary emphasizing the importance of owning your own memory harness in AI applications rather than relying on third-party API-based

MiniMax M2.7 is now on Together AI. Trained by letting it run its own RL loop, resulting in the highest open-source score on MLE Bench Lite.

ToolsDGX agent

MiniMax M2.7 is now available on Together AI, offering a powerful open-source model trained using a self-directed reinforcement learning loop. This training methodology enabled the model to achieve th

Qwen 2511 fp8 mixed taking 30–40s per image edit — which GGUF should I use?

Model ReleasesDGX agent

This r/StableDiffusion thread addresses performance issues with the Qwen-Image-Edit-2511 FP8 mixed model in ComfyUI, where users experience slow image edit times of 30–40 seconds per step (or overall)

11 Apr 2026

My 2026 Ollama Setup Guide: What Actually Works Best for Daily Use on Consumer Hardware

Local AiDGX agent

This r/ollama community post is a practitioner's guide sharing personal, real-world experience running Ollama on everyday consumer hardware in 2026, covering which models, quantization settings, and c

We've been tricked, again. Many of the thousands of bugs and vulnerabilities Mythos found are in older software are impossible to exploit. A…

Model ReleasesDGX agent

We've been tricked, again. Many of the thousands of bugs and vulnerabilities Mythos found are in older software are impossible to exploit. And the severe zero-day reports rely on just 198 manual revie

10 Apr 2026

A Parameter-Efficient Transfer Learning Approach through Multitask Prompt Distillation and Decomposition for Clinical NLP

Model ReleasesDGX agent

arXiv:2604.06650v1 Announce Type: cross Abstract: Existing prompt-based fine-tuning methods typically learn task-specific prompts independently, imposing significant computing and storage overhead at

A playable, digital piano built with Gemma 4 E2B (2B). https://x.com/0xCVYH/status/2039929097141318129?s=20

Model ReleasesDGX agent

A playable, digital piano built with Gemma 4 E2B (2B). https://x.com/0xCVYH/status/2039929097141318129?s=20 Google lancou 'Agent Skills' junto com o Gemma 4 Um app Android onde voce importa skills e o

Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference

Model ReleasesDGX agent

arXiv:2604.08133v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become a dominant architecture for scaling large language models due to their sparse activation mechanism. However, the s

Auditing Black-Box LLM APIs with a Rank-Based Uniformity Test

Local AiDGX agent

arXiv:2506.06975v5 Announce Type: replace-cross Abstract: As API access becomes a primary interface to large language models (LLMs), users often interact with black-box systems that offer little trans

Bayesian E(3)-Equivariant Interatomic Potential with Iterative Restratification of Many-body Message Passing

Model ReleasesDGX agent

arXiv:2510.03046v2 Announce Type: replace Abstract: Machine learning potentials (MLPs) have become essential for large-scale atomistic simulations, enabling ab initio-level accuracy with computational

Beyond Case Law: Evaluating Structure-Aware Retrieval and Safety in Statute-Centric Legal QA

Model ReleasesDGX agent

arXiv:2604.06173v1 Announce Type: cross Abstract: Legal QA benchmarks have predominantly focused on case law, overlooking the unique challenges of statute-centric regulatory reasoning. In statutory do

Cram Less to Fit More: Training Data Pruning Improves Memorization of Facts

ResearchDGX agent

arXiv:2604.08519v1 Announce Type: new Abstract: Large language models (LLMs) can struggle to memorize factual knowledge in their parameters, often leading to hallucinations and poor performance on kno

Create Expert Content: Local Testing of a Multi-Agent System with Memory

Model ReleasesDGX agent

In support of our mission to accelerate the developer journey on Google Cloud, we built Dev Signal: a multi-agent system designed to transform raw community signals into reliable technical guidance by

CycleChart: A Unified Consistency-Based Learning Framework for Bidirectional Chart Understanding and Generation

Model ReleasesDGX agent

arXiv:2512.19173v2 Announce Type: replace Abstract: Current chart-related tasks, such as chart generation (NL2Chart), chart schema parsing, chart data parsing, and chart question answering (ChartQA),

Distilling Specialized Orders for Visual Generation

ResearchDGX agent

arXiv:2504.17069v2 Announce Type: replace Abstract: Autoregressive (AR) image generators are becoming increasingly popular due to their ability to produce high-quality images and their scalability. Ty

Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents

Model ReleasesDGX agent

arXiv:2604.08369v1 Announce Type: cross Abstract: Inference-time compute scaling has emerged as a powerful technique for improving the reliability of large language model (LLM) agents, but existing me

DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning

Model ReleasesDGX agent

arXiv:2410.17473v2 Announce Type: replace Abstract: In reinforcement learning (RL), temporal difference (TD) error is known to be related to the firing rate of dopamine neurons. It has been observed t

Dual-level Modality Debiasing Learning for Unsupervised Visible-Infrared Person Re-Identification

Model ReleasesDGX agent

arXiv:2512.03745v2 Announce Type: replace Abstract: Two-stage learning pipeline has achieved promising results in unsupervised visible-infrared person re-identification (USL-VI-ReID). It first perform

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions

Model ReleasesDGX agent

arXiv:2604.07772v1 Announce Type: new Abstract: Open-world video anomaly detection (OWVAD) aims to detect and explain abnormal events under different anomaly definitions, which is important for applic

Evaluating LLM-Based 0-to-1 Software Generation in End-to-End CLI Tool Scenarios

Model ReleasesDGX agent

arXiv:2604.06742v1 Announce Type: cross Abstract: Large Language Models (LLMs) are driving a shift towards intent-driven development, where agents build complete software from scratch. However, existi

Face2Scene: Using Facial Degradation as an Oracle for Diffusion-Based Scene Restoration

TutorialsDGX agent

arXiv:2603.16570v2 Announce Type: replace Abstract: Recent advances in image restoration have enabled high-fidelity recovery of faces from degraded inputs using reference-based face restoration models

FLeX: Fourier-based Low-rank EXpansion for multilingual transfer

Model ReleasesDGX agent

arXiv:2604.06253v1 Announce Type: cross Abstract: Cross-lingual code generation is critical in enterprise environments where multiple programming languages coexist. However, fine-tuning large language

From Ground Truth to Measurement: A Statistical Framework for Human Labeling

SafetyDGX agent

arXiv:2604.07591v1 Announce Type: cross Abstract: Supervised machine learning assumes that labeled data provide accurate measurements of the concepts models are meant to learn. Yet in practice, human

FVD: Inference-Time Alignment of Diffusion Models via Fleming-Viot Resampling

SafetyDGX agent

arXiv:2604.06779v1 Announce Type: new Abstract: We introduce Fleming-Viot Diffusion (FVD), an inference-time alignment method that resolves the diversity collapse commonly observed in Sequential Monte

HiCI: Hierarchical Construction-Integration for Long-Context Attention

Model ReleasesDGX agent

arXiv:2603.20843v2 Announce Type: replace Abstract: Long-context language modeling is commonly framed as a scalability challenge of token-level attention, yet local-to-global information structuring r

How to sketch a learning algorithm

TutorialsDGX agent

arXiv:2604.07328v1 Announce Type: new Abstract: How does the choice of training data influence an AI model? This question is of central importance to interpretability, privacy, and basic science. At i

In-Context Decision Making for Optimizing Complex AutoML Pipelines

Model ReleasesDGX agent

arXiv:2508.13657v2 Announce Type: replace-cross Abstract: Combined Algorithm Selection and Hyperparameter Optimization (CASH) has been fundamental to traditional AutoML systems. However, with the adva

KEO: Knowledge Extraction on OMIn via Knowledge Graphs and RAG for Safety-Critical Aviation Maintenance

Model ReleasesDGX agent

arXiv:2510.05524v2 Announce Type: replace Abstract: We present Knowledge Extraction on OMIn (KEO), a domain-specific knowledge extraction and reasoning framework with large language models (LLMs) in s

Mitigating Spurious Background Bias in Multimedia Recognition with Disentangled Concept Bottlenecks

Model ReleasesDGX agent

arXiv:2510.15770v3 Announce Type: replace Abstract: Concept Bottleneck Models (CBMs) enhance interpretability by predicting human-understandable concepts as intermediate representations. However, exis

Noise Immunity in In-Context Tabular Learning: An Empirical Robustness Analysis of TabPFN's Attention Mechanisms

Model ReleasesDGX agent

arXiv:2604.04868v2 Announce Type: replace-cross Abstract: Tabular foundation models (TFMs) such as TabPFN (Tabular Prior-Data Fitted Network) are designed to generalize across heterogeneous tabular da

On the Step Length Confounding in LLM Reasoning Data Selection

ResearchDGX agent

arXiv:2604.06834v1 Announce Type: cross Abstract: Large reasoning models have recently demonstrated strong performance on complex tasks that require long chain-of-thought reasoning, through supervised

Quality-preserving Model for Electronics Production Quality Tests Reduction

SafetyDGX agent

arXiv:2604.06451v1 Announce Type: new Abstract: Manufacturing test flows in high-volume electronics production are typically fixed during product development and executed unchanged on every unit, even

Qwen 3.6 Plus on AI Gateway

Model ReleasesDGX agent

Alibaba's Qwen 3.6 Plus is now available on Vercel's AI Gateway, accessible via a unified API with no additional provider accounts required. Compared to Qwen 3.5 Plus, this model adds stronger age...

qwen3.5 no se instala

Local AiDGX agent

Users in the r/ollama community reported that Qwen3.5 models fail to install or load in Ollama, with common errors including 'Error: 500 Internal Server Error: unable to load model' even after a ...

Scaling-Aware Data Selection for End-to-End Autonomous Driving Systems

AgentsDGX agent

arXiv:2604.08366v1 Announce Type: cross Abstract: Large-scale deep learning models for physical AI applications depend on diverse training data collection efforts. These models and correspondingly, th

Sell More, Play Less: Benchmarking LLM Realistic Selling Skill

Model ReleasesDGX agent

arXiv:2604.07054v2 Announce Type: replace Abstract: Sales dialogues require multi-turn, goal-directed persuasion under asymmetric incentives, which makes them a challenging setting for large language

SpatialMosaic: A Multiview VLM Dataset for Partial Visibility

Model ReleasesDGX agent

arXiv:2512.23365v3 Announce Type: replace Abstract: The rapid progress of Multimodal Large Language Models (MLLMs) has unlocked the potential for enhanced 3D scene understanding and spatial reasoning.

Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation

Model ReleasesDGX agent

arXiv:2604.07028v1 Announce Type: cross Abstract: Strategic interaction in adversarial domains such as law, diplomacy, and negotiation is mediated by language, yet most game-theoretic models abstract

SubSearch: Intermediate Rewards for Unsupervised Guided Reasoning in Complex Retrieval

AgentsDGX agent

arXiv:2604.07415v1 Announce Type: cross Abstract: Large language models (LLMs) are probabilistic in nature and perform more reliably when augmented with external information. As complex queries often

Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding

Model ReleasesDGX agent

arXiv:2604.07753v1 Announce Type: cross Abstract: Empowering Large Multimodal Models (LMMs) with image generation often leads to catastrophic forgetting in understanding tasks due to severe gradient c

Tensor-Augmented Convolutional Neural Networks: Enhancing Expressivity with Generic Tensor Kernels

Model ReleasesDGX agent

arXiv:2604.08072v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) excel at extracting local features hierarchically, but their performance in capturing complex correlations hinges h

The Download: an exclusive Jeff VanderMeer story and AI models too scary to release

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Constellations —Constellations is a short story by Jeff Vander

The Stepwise Informativeness Assumption: Why are Entropy Dynamics and Reasoning Correlated in LLMs?

Model ReleasesDGX agent

arXiv:2604.06192v1 Announce Type: cross Abstract: Recent work uses entropy-based signals at multiple representation levels to study reasoning in large language models, but the field remains largely em

Those of you that run Openclaw with Ollama Pro, do you need the local ollama to use cloud?

Local AiDGX agent

When using Ollama Pro's cloud models with OpenClaw, a local Ollama installation is not strictly required — OpenClaw is an AI agent execution layer that handles tools, memory, scheduling, and messa...

Transformer See, Transformer Do: Copying as an Intermediate Step in Learning Analogical Reasoning

ResearchDGX agent

arXiv:2604.06501v1 Announce Type: new Abstract: Analogical reasoning is a hallmark of human intelligence, enabling us to solve new problems by transferring knowledge from one situation to another. Yet

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation

Model ReleasesDGX agent

arXiv:2604.07894v1 Announce Type: new Abstract: Personalized large language models (PLLMs) have garnered significant attention for their ability to align outputs with individual's needs and preference

VisCoder2: Building Multi-Language Visualization Coding Agents

Model ReleasesDGX agent

arXiv:2510.23642v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently enabled coding agents capable of generating, executing, and revising visualization code. However, e

9 Apr 2026

Hermes will soon be a serious creative tool. More on this soon!

ResearchDGX agent

Nous Research's Hermes model line has been deliberately developed with a strong emphasis on creative capability. The training dataset for Hermes 4 expanded to 50x more data tokens than Hermes 3, w...

I looked at their prompts, It's complete bs They are literally providing all of the insight to the LLM upfront > Are there any security vuln…

ResearchDGX agent

I looked at their prompts, It's complete bs They are literally providing all of the insight to the LLM upfront > Are there any security vulnerabilities in this code? Consider the behavior of the SEQ_L

if you prefer blog format: https://blog.langchain.com/deep-agents-deploy-an-open-alternative-to-claude-managed-agents/

Model ReleasesDGX agent

LangChain launched **Deep Agents Deploy** in beta as an open-source, model-agnostic alternative to Anthropic's Claude Managed Agents, allowing developers to deploy production-ready agents via a sin...

Parax: Parametric Modeling in JAX + Equinox [P]

ResearchDGX agent

**Paramax** (also referred to as 'Parax' in the Reddit post title) is a small Python library by Daniel Ward that provides parameterizations and parameter constraints for JAX PyTrees, designed to wo...

v0.20.5-rc1

Local AiDGX agent

<channel|>This release update for `local-ai` (version v0.20.5-rc1) expands model compatibility by integrating several popular large language models. It allows users to run models such as Kimi-K2.5,...

8 Apr 2026

'But here is what we found when we tested: We took the specific vulnerabilities Anthropic showcases in their announcement, isolated the rele…

ResearchDGX agent

'But here is what we found when we tested: We took the specific vulnerabilities Anthropic showcases in their announcement, isolated the relevant code, and ran them through small, cheap, open-weights m

← Previous
1…292293294295296…1035
Next →