AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
10 Apr 2026

LTX 2.3 - Image + Audio + Video ControlNet (IC-LoRA) to Video

Local AiDGX agent

LTX-2.3 is a DiT-based audio-video foundation model from Lightricks that generates synchronized video and audio within a single model pass, representing a significant upgrade over LTX-2 with improv...

Mac mini M4 48GB

Local AiDGX agent

The Mac Mini M4 Pro with 48GB unified memory is a popular choice in the local AI community for running large language models via Ollama, as its Apple Silicon architecture makes all 48GB of RAM dire...

MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstracts

Model ReleasesDGX agent

arXiv:2604.06505v1 Announce Type: cross Abstract: Large language models (LLMs) are widely explored for reasoning-intensive research tasks, yet resources for testing whether they can infer scientific c

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Novel View Synthesis as Video Completion

ResearchDGX agent

arXiv:2604.08500v1 Announce Type: new Abstract: We tackle the problem of sparse novel view synthesis (NVS) using video diffusion models; given K (approx 5) multi-view images of a scene and their

Open-Ended Instruction Realization with LLM-Enabled Multi-Planner Scheduling in Autonomous Vehicles

Model ReleasesDGX agent

arXiv:2604.08031v1 Announce Type: cross Abstract: Most Human-Machine Interaction (HMI) research overlooks the maneuvering needs of passengers in autonomous driving (AD). Natural language offers an int

Our first successful Gemma 4 Runtime in London with @swyx @patloeber @nick_kango @cormacb and others! 💎Great to go out for a run and talk a…

Model ReleasesDGX agent

Our first successful Gemma 4 Runtime in London with @swyx @patloeber @nick_kango @cormacb and others! 💎Great to go out for a run and talk about Gemma, agents, evals and more @osanseviero @swyx and oth

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory

Model ReleasesDGX agent

arXiv:2604.08000v1 Announce Type: cross Abstract: Proactivity is a core expectation for AGI. Prior work remains largely confined to laboratory settings, leaving a clear gap in real-world proactive age

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing

Model ReleasesDGX agent

arXiv:2604.07230v2 Announce Type: replace Abstract: Achieving physically accurate object manipulation in image editing is essential for its potential applications in interactive world models. However,

🚀 Qwen Code v0.14.0 – v0.14.2 are now available Channels:Control Qwen Code remotely from Telegram, DingTalk, or WeChat — send a message fro…

Model ReleasesDGX agent

🚀 Qwen Code v0.14.0 – v0.14.2 are now available Channels:Control Qwen Code remotely from Telegram, DingTalk, or WeChat — send a message from your phone, get results on your server Cron Jobs :Schedule

Reasoning Fails Where Step Flow Breaks

ResearchDGX agent

arXiv:2604.06695v1 Announce Type: new Abstract: Large reasoning models (LRMs) that generate long chains of thought now perform well on multi-step math, science, and coding tasks. However, their behavi

REVEAL: Reasoning-Enhanced Forensic Evidence Analysis for Explainable AI-Generated Image Detection

Model ReleasesDGX agent

arXiv:2511.23158v2 Announce Type: replace-cross Abstract: The rapid progress of visual generative models has made AI-generated images increasingly difficult to distinguish from authentic ones, posing

RewardFlow: Generate Images by Optimizing What You Reward

SafetyDGX agent

arXiv:2604.08536v1 Announce Type: new Abstract: We introduce RewardFlow, an inversion-free framework that steers pretrained diffusion and flow-matching models at inference time through multi-reward La

SearchAD: Large-Scale Rare Image Retrieval Dataset for Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.08008v1 Announce Type: new Abstract: Retrieving rare and safety-critical driving scenarios from large-scale datasets is essential for building robust autonomous driving (AD) systems. As dat

SpecQuant: Spectral Decomposition and Adaptive Truncation for Ultra-Low-Bit LLMs Quantization

Model ReleasesDGX agent

arXiv:2511.11663v2 Announce Type: replace-cross Abstract: The emergence of accurate open large language models (LLMs) has sparked a push for advanced quantization techniques to enable efficient deploy

SVGFusion: A VAE-Diffusion Transformer for Vector Graphic Generation

ResearchDGX agent

arXiv:2412.10437v3 Announce Type: replace Abstract: Generating high-quality Scalable Vector Graphics (SVGs) from text remains a significant challenge. Existing LLM-based models that generate SVG code

Visually-grounded Humanoid Agents

Model ReleasesDGX agent

arXiv:2604.08509v1 Announce Type: new Abstract: Digital human generation has been studied for decades and supports a wide range of real-world applications. However, most existing systems are passively

What Your Local LLM Actually Sees: Debugging Ollama Traffic in Quarkus with mitmproxy

Local AiDGX agent

This tutorial demonstrates how to use mitmproxy to inspect the actual HTTP traffic sent from a Quarkus application to a local Ollama model via its OpenAI-compatible endpoint, revealing the real JSO...

When Personalization Tricks Detectors: The Feature-Inversion Trap in Machine-Generated Text Detection

Model ReleasesDGX agent

arXiv:2510.12476v2 Announce Type: replace Abstract: Large language models (LLMs) have grown more powerful in language generation, producing fluent text and even imitating personal style. Yet, this abi

Yet another illustration of why LLMs aren’t even close to being AGI.

Model ReleasesDGX agent

Yet another illustration of why LLMs aren’t even close to being AGI. The world’s best LLMs are still terrible at poker. We put each model into a 200bb heads-up NLHE match against GTO Wizard AI. The be

9 Apr 2026

A developer’s guide to architecting reliable GPU infrastructure at scale

Model ReleasesDGX agent

Editor’s note: This blog post outlines Google Cloud’s GPU AI/ML infrastructure reliability strategy, and will be updated with links to new community articles as they appear. As we enter the era of mul

Tried running LLMs locally to save API costs… ended up waiting 13 minutes for ONE response 🤡

Local AiDGX agent

A Reddit post in r/ollama describes a user's experience attempting to run LLMs locally via Ollama to avoid cloud API costs, only to encounter severely degraded performance — waiting 13 minutes for ...

8 Apr 2026

I’ve uploaded a new paper on arXiv (co-authored by @rasbt): MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning In Parameter-Efficient…

Model ReleasesDGX agent

I’ve uploaded a new paper on arXiv (co-authored by @rasbt): MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning In Parameter-Efficient Fine-Tuning, a key question may not just be how low-rank th

Wordle 1,754 4/6 ⬛🟨⬛⬛🟨 🟨🟨🟨⬛⬛ 🟩🟩🟨🟩🟨 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

I was unable to retrieve the specific content of the linked X (Twitter) post from Anthropic at that URL, and the web search did not return results directly related to it. The post ID (2042020348451...

7 Apr 2026

GLM-5.1 is live everywhere you use the Kilo Gateway (VS Code extension, Cloud Agents, KiloClaw, etc). Thank you @Zai_org! ⚡️

Model ReleasesDGX agent

Z.AI's GLM-5.1, a next-generation flagship model for agentic engineering released in April 2026, is now available across all Kilo Code surfaces — including the VS Code extension, Cloud Agents, and ...

Mythos is very powerful, and should feel terrifying. I am proud of our approach to responsibly preview it with cyber defenders, rather than …

Model ReleasesDGX agent

Mythos is very powerful, and should feel terrifying. I am proud of our approach to responsibly preview it with cyber defenders, rather than generally releasing it into the wild. Model card here: https

19 Aug 2026

Aloha! 🌺Introducing Ornith-1.5, a family of open-source LLMs spanning 9B Dense, 35B MoE, and 397B MoE, trained with self-improving strategi…

Model ReleasesDGX agent

Aloha! 🌺Introducing Ornith-1.5, a family of open-source LLMs spanning 9B Dense, 35B MoE, and 397B MoE, trained with self-improving strategies. It achieves state-of-the-art performance among open-sourc

Am I doing something wrong? Qwen 3.8 27B seems useless for agentic coding

Model ReleasesDGX agent

I have been using local models on/off for like 2 years or so but never really used them extensively because the closed ones were always much better. Once Qwen 3.8 27B was released I decided to give it

Beyond the Trace: Coupling an Interpretable Reasoning-State Readout to Native MoE Routing

SafetyDGX agent

arXiv:2608.17638v1 Announce Type: new Abstract: What a reasoning model writes is only a partial record of the process that produces it. We introduce a two-level internal readout for mixture-of-experts

I am so tired of the PR.

Model ReleasesDGX agent

I am so tired of the PR. How Anthropic's new results post would read without the PR: Claude orchestrated open-source protein design models, PXDesign, RFdiffusion, Genie, BoltzGen, from a 30k-token exp

Improving Complex Moire Removal with Generative Supervision

Model ReleasesDGX agent

arXiv:2608.17883v1 Announce Type: new Abstract: The availability of high-quality paired data is essential for training learning-based image demoireing models. However, it remains challenging for exist

Key-Frame Reasoning with SAM3: Third Place Solution for the MeViS-Text Track of the 8th LSVOS Challenge

Model ReleasesDGX agent

arXiv:2608.17279v1 Announce Type: new Abstract: This report presents a two-stage, training-free solution for the MeViS-Text track of the 8th LSVOS Challenge. The task requires a model to localize and

Leveraging existing sparse point annotations for benthic imagery dense segmentation

Model ReleasesDGX agent

arXiv:2608.17561v1 Announce Type: new Abstract: The health of marine ecosystems is a critical indicator of global environmental change, yet the physical constraints of underwater observation and the i

MANIGUARD: A Benchmark and Data Suite for Specification-Grounded Safety Evaluation and Improvement of Robotic Manipulation

Model ReleasesDGX agent

arXiv:2608.17386v1 Announce Type: new Abstract: Foundation-model policies for robotic manipulation are advancing rapidly on task success, but rigorous evaluation of whether they succeed safely is stil

Mixture-of-Expert Blocks Contain Strong Hallucination Detection Signals

Local AiDGX agent

arXiv:2608.17687v1 Announce Type: new Abstract: Despite their widespread use, Large Language Models (LLMs) remain limited by a fundamental problem: the generation of plausible but false content, known

Policy-Invariant Reward Shaping from LLM Feedback: A Framework for Hybrid RL Agents

SafetyDGX agent

arXiv:2608.18008v1 Announce Type: cross Abstract: Combining large language models with reinforcement learning is increasingly explored, yet the theoretical status of LLM-derived reward signals is ofte

Probing the Prefill: Detecting Code Vulnerabilities via Latent Activations

Model ReleasesDGX agent

arXiv:2608.16970v1 Announce Type: cross Abstract: LLM-based code generation is now embedded in mission-critical pipelines, but defenses against vulnerable output remain post-hoc -- static analyzers, f

PTXBench: Benchmark and Adapt LLMs for GPU Kernel Optimization with Architecture-specific PTX

Model ReleasesDGX agent

arXiv:2608.17379v1 Announce Type: cross Abstract: We introduce PTXBench, a benchmark for evaluating and adapting large language models (LLMs) to use architecture-specific PTX for GPU kernel optimizati

Q-Interference: Memory-Efficient Phase-Aware Quantum-Inspired Attention

Model ReleasesDGX agent

arXiv:2608.17288v1 Announce Type: new Abstract: GPT attention measures token compatibility through dot-product similarity. This mechanism is simple, effective, and memory-efficient. But it does not ex

S^3AM: A Single-Stream SAM with Reliability-Calibrated Frequency Adapter for Multi-modal Salient Object Detection

Model ReleasesDGX agent

arXiv:2608.17475v1 Announce Type: new Abstract: Vision foundation models have recently advanced multi-modal salient object detection (MSOD) through parameter-efficient tuning and prompt learning. Howe

Teach and Grow: An Agent-Centered Architecture for General Robot Learning

SafetyDGX agent

arXiv:2608.17209v1 Announce Type: cross Abstract: End-to-end vision-language-action (VLA) and world-action models offer an elegant route to general-purpose robotics, but their reliability is bounded b

Towards Safer RAG: Only Agents Capable of System 2 Thinking may Access Untrusted Documents

ResearchDGX agent

arXiv:2608.17153v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has significantly enhanced the performance of large language models (LLMs), yet these systems remain vulnerable to

18 Aug 2026

A Large-Scale Chinese Knowledge Graph-Text Alignment Dataset for Benchmarking Knowledge-Grounded LLMs

Model ReleasesDGX agent

arXiv:2510.06039v2 Announce Type: replace-cross Abstract: Reliable evaluation of knowledge-grounded Large Language Models (LLMs) in Chinese requires resources that explicitly align Chinese-language te

A Unified DINOv2-Based Framework for LVEF Estimation, GLS Dysfunction Classification, and Early Cardiotoxicity Prediction

Model ReleasesDGX agent

arXiv:2608.14750v1 Announce Type: cross Abstract: Left ventricular ejection fraction (LVEF) estimation (Task 1), global longitu-dinal strain (GLS)-based dysfunction classification (Task 2), and early

Aborted but Not Forgotten: KV-Cache Retention Breaks Rollback Consistency in Language Agents

Model ReleasesDGX agent

arXiv:2608.15939v1 Announce Type: new Abstract: Stateful language agents assume a rejected branch can be taken back by clearing it from the application transcript. We show this breaks when the serving

Benchmarking Identity-Sensitive LLM Outputs for Surveillance and Security Robots

Model ReleasesDGX agent

arXiv:2608.16030v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to generate textual robot design specifications, interaction policies, and risk assessments during ea

Beyond Asking: A Pipeline for Personalized Game Generation that Reads Players from Behavior

Model ReleasesDGX agent

arXiv:2608.16196v1 Announce Type: new Abstract: Personalized game generation requires inferring a player's abilities and behavioral style from how they play. Large language models have made this infer

Characterization of Thermal Systems from Noisy and Low-resolution Measurements Using Dynamic Mode Decomposition

Model ReleasesDGX agent

arXiv:2608.14581v1 Announce Type: cross Abstract: Thermal monitoring in practical applications is often constrained by sparse sensing, measurement noise, and limited spatial resolution, which hinder t

Decorrelation Is Not Complementarity: Skill, Not Lineage, Governs Trusted-Monitor Ensembles

ResearchDGX agent

arXiv:2608.16190v1 Announce Type: cross Abstract: Trusted monitoring has a cheap, trusted model score a stronger untrusted model's actions, and a diverse ensemble of them beats a single stronger monit

DeepInsight II: One Trace from Benchmark to Robot

Model ReleasesDGX agent

arXiv:2608.16556v1 Announce Type: new Abstract: Across a Physical AI stack, evaluation maturity is inversely aligned with deployment risk: foundation models enjoy mature, standardized harnesses, while

Defake-o3: From Speculative Rationales to Verifiable Evidence for Explainable AIGI Detection

Model ReleasesDGX agent

arXiv:2608.16259v1 Announce Type: cross Abstract: The rapid progress of image generation models calls for AI-generated image (AIGI) detectors that are not only accurate but also explainable and reliab

DepthArb: Training-Free Depth-Arbitrated Generation for Occlusion-Robust Image Synthesis

Model ReleasesDGX agent

arXiv:2603.23924v2 Announce Type: replace Abstract: Text-to-image models often struggle to synthesize correct occlusion relationships among multiple objects, especially in densely overlapping regions.

Do LLMs Know What to Ask and When? Evaluating Multi-Turn Information Seeking

ResearchDGX agent

arXiv:2608.14808v1 Announce Type: new Abstract: When a user question is underspecified, a capable model should recognize that its context is insufficient, identify the missing information, ask for it,

Eigenanalysis framework for autoregressive neural emulators of multi-scale chaotic dynamics

SafetyDGX agent

arXiv:2608.16084v1 Announce Type: new Abstract: Neural autoregressive models have rapidly emerged as powerful emulators of high-dimensional chaotic systems, yet their long-term instability and error g

From LLM Inference to Agentic Workloads: Characterization and Implications for Serving Systems

Model ReleasesDGX agent

arXiv:2608.15127v1 Announce Type: cross Abstract: Agentic applications are shifting AI serving from isolated model inference to long-running workloads in which LLMs coordinate tools, environments, and

FusionBERT: Multi-View Image--3D Retrieval via Cross-Attention Visual Fusion and Normal-Aware 3D Encoder

SafetyDGX agent

arXiv:2604.02583v2 Announce Type: replace Abstract: We propose FusionBERT, a novel multi-view visual fusion framework for image--3D multimodal retrieval. Existing image--3D representation learning met

Governance at the Boundary: How Agent Decomposition Degrades Policy Compliance

Model ReleasesDGX agent

arXiv:2608.16055v1 Announce Type: new Abstract: Existing agent benchmarks ask whether the agent finished the task. We ask whether it finished it within policy. We introduce Fiducia-bench, a benchmark

HarnessEval-W: Agentifying the Evaluation of Visual Worlds

Model ReleasesDGX agent

arXiv:2608.16859v1 Announce Type: new Abstract: A benchmark should deliver more than a scalar score: what makes an evaluation trustworthy is the reasoning that justifies the score. This is especially

Incoherent by Design? On the Moral Self-Consistency of LLMs

Model ReleasesDGX agent

arXiv:2608.15354v1 Announce Type: new Abstract: LLMs are increasingly used in morally sensitive contexts, yet it is unclear whether they apply ethical principles consistently across situations. A mode

INSPIRE: A Benchmark for Instruction-Aware Speech Retrieval

Model ReleasesDGX agent

arXiv:2608.16203v1 Announce Type: cross Abstract: Existing speech retrieval systems rely on fixed similarity matching and cannot adapt to diverse user intents. We introduce INSPIRE, the first benchmar

LlamaRec-LKG-RAG: A Single-Pass, Learnable Knowledge Graph-RAG Framework for LLM-Based Ranking

Model ReleasesDGX agent

arXiv:2506.07449v2 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have driven their adoption in recommender systems through Retrieval-Augmented Generation (RAG)

← Previous
1…324325326327328…1042
Next →