AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
8 Apr 2026

If you’ve encountered garbled output like this while using GLM-5 or GLM-5.1 on our official service, the issue is now resolved. We've patche…

Model ReleasesDGX agent

If you’ve encountered garbled output like this while using GLM-5 or GLM-5.1 on our official service, the issue is now resolved. We've patched the underlying inference-side bugs and will be releasing a

it all makes sense now. dario was still at openai in 2019. he left next year and took his marketing playbook with him. hasn't changed a thin…

ResearchDGX agent

A crypto/DeFi developer named banteg posted on X (Twitter) observing that Dario Amodei was still at OpenAI in 2019, where he worked under Sam Altman from 2016 to 2020 as VP of Research, playing an...

Mornings feel different lately. Before I even get organized, I already have 5+ agents running. Same thing before I go to sleep. Claude Code,…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

Mornings feel different lately. Before I even get organized, I already have 5+ agents running. Same thing before I go to sleep. Claude Code, ChatGPT, Qodo, Notion, Linear, NotebookLM. All moving in pa

New GKE Cloud Storage FUSE Profiles take the guesswork out of configuring AI storage

TutorialsDGX agent

In the world of AI/ML, data is the fuel that drives training and inference workloads. For Google Kubernetes Engine (GKE) users, Cloud Storage FUSE provides high-performance, scalable access to data st

OpenClaw 2026.4.7 🦞 🔮 openclaw infer 🎬 music + video editing 💾 session branch/restore 🔗 webhook-driven TaskFlows 🤖 Arcee, Gemma 4, Oll…

Model ReleasesDGX agent

OpenClaw 2026.4.7 🦞 🔮 openclaw infer 🎬 music + video editing 💾 session branch/restore 🔗 webhook-driven TaskFlows 🤖 Arcee, Gemma 4, Ollama vision 🧠 memory-wiki: persistent knowledge, not just vibes Bec

Real-time inference for robots at Physical Intelligence

ToolsDGX agent

Physical Intelligence (Pi) is building a general-purpose robotic intelligence system whose core Visual-Language-Action (VLA) model takes visual observations, natural-language instructions, and the...

Safetensors is Joining the PyTorch Foundation

ToolsDGX agent

Safetensors, a secure tensor serialization format originally developed by Hugging Face, has joined the PyTorch Foundation as a foundation-hosted project under the Linux Foundation, alongside DeepSp...

Strong release! GLM-5.1 is a DeepSeek-V3.2-like architecture (including MLA and DeepSeek Sparse Attention) but with more layers. And the ben…

Model ReleasesDGX agent

Strong release! GLM-5.1 is a DeepSeek-V3.2-like architecture (including MLA and DeepSeek Sparse Attention) but with more layers. And the benchmarks look better throughout! Looks like THE flagship open

The feature I most want from AI labs right now is documentation on which underlying search engines they use when their chat tools run a sear…

Model ReleasesDGX agent

The feature I most want from AI labs right now is documentation on which underlying search engines they use when their chat tools run a search OpenAI and Anthropic and Meta AI all have search and I ha

The new Anthropic managed agents API is basically the Letta API that we've had since a year ago, but closed source and with provider lock-in…

Model ReleasesDGX agent

The new Anthropic managed agents API is basically the Letta API that we've had since a year ago, but closed source and with provider lock-in. They even have read-only memory blocks and memory block sh

We partnered with @Zai_org to bring GLM-5.1 to Modal. Free to try as an endpoint for the next month. GLM-5.1 further improves upon GLM-5's c…

Model ReleasesDGX agent

We partnered with @Zai_org to bring GLM-5.1 to Modal. Free to try as an endpoint for the next month. GLM-5.1 further improves upon GLM-5's coding abilities and long-horizon effectiveness. Introducing

7 Apr 2026

got opencode working in taugentic, so now i can use glm-5.1 with my coding plan.

Model ReleasesDGX agent

User Kevin Kern reported successfully integrating OpenCode into Taugentic (an AI-powered browser/agent environment), enabling use of GLM-5.1 via Z.ai's GLM Coding Plan. GLM-5.1 is Z.ai's next-gener...

How can you improve your agentic search pipeline? I just wrote a blog post with @tech_optimist from @lancedb to answer exactly that. TLDR: -…

Model ReleasesDGX agent

How can you improve your agentic search pipeline? I just wrote a blog post with @tech_optimist from @lancedb to answer exactly that. TLDR: - Parse files and take page-level screenshots with LiteParse,

I'm a big fan of the pelican GLM-5.1 drew me today, it even animated it! https://simonwillison.net/2026/Apr/7/glm-51/

ToolsDGX agent

Simon Willison tested Z.ai's GLM-5.1 using his standard 'pelican on a bicycle' SVG benchmark, and the model stood out by spontaneously generating a full HTML page with both the SVG and a separate ...

Join the ARC Prize team -- help us build ARC-AGI-4 and ARC-AGI-5

Model ReleasesDGX agent

Join the ARC Prize team -- help us build ARC-AGI-4 and ARC-AGI-5 Platform Engineer - Benchmark Lead ARC Prize Foundation is hiring a senior engineer to build our benchmark platform * Expand ARC-AGI-3

Oh no.

ApplicationsDGX agent

I was unable to retrieve the specific tweet at the URL provided (`https://x.com/emollick/status/2041600435320959330`). The tweet ID `2041600435320959330` does not appear in any search results, and ...

Technically all the vulnerabilities in public facing code are in the training data Make of that what you will

IndustryDGX agent

Technically all the vulnerabilities in public facing code are in the training data Make of that what you will Mythos Preview has already found thousands of high-severity vulnerabilities—including some

v0.20.4-rc2: gemma4: Disable FA on older GPUs where it doesn't work (#15403)

Local AiDGX agent

Ollama v0.20.4-rc2 is a release candidate that addresses a compatibility issue with Flash Attention (FA) for the Gemma 4 model on older GPUs. CUDA versions older than 7.5 lack the support needed t...

Visually rich documents are especially challenging for agents. Tables, charts, and images often break traditional document pipelines, making…

Model ReleasesDGX agent

Visually rich documents are especially challenging for agents. Tables, charts, and images often break traditional document pipelines, making complex reasoning difficult📄 So we teamed up with @lancedb

24 Aug 2026

A Hybrid Edge Cloud Digital Twin for Welfare-Constrained Control in Poultry Production

Local AiDGX agent

arXiv:2608.20367v1 Announce Type: cross Abstract: Poultry production operates under tightly coupled environmental and biological dynamics, yet commercial climate control remains largely heuristic, lim

Affective Context Amplifies Sycophancy in LLM Responses

ResearchDGX agent

arXiv:2608.21242v1 Announce Type: new Abstract: As conversational companions, large language models (LLMs) often have access to users' emotional states. We study how this affective context modulates L

AffordAny: Open-World 3D Affordance Grounding from Monocular RGB Images via Vision-Language-Guided Geometric Reasoning

Model ReleasesDGX agent

arXiv:2608.20720v1 Announce Type: new Abstract: Open-world 3D affordance grounding requires localizing functional object parts in 3D given free-form language queries. Existing methods typically assume

Aggregating Visual Information with Optimal Transport for VideoLM Token Compression

ResearchDGX agent

arXiv:2608.20473v1 Announce Type: new Abstract: Video language models process videos as dense visual-token sequences with substantial representational redundancy. Compressing these sequences is theref

ARQ: Agentic CodeQL Query Refinement for C/C++ Vulnerability Detection

Model ReleasesDGX agent

arXiv:2608.20637v1 Announce Type: cross Abstract: Static analyzers have been widely adopted for vulnerability detection in C/C++ programs. Query-based static analyzers (e.g., CodeQL) encode vulnerable

Bankruptcy Prediction via Hybrid Resampling and Stacking Ensemble Techniques with Explainable Artificial Intelligence (XAI)-Driven Analysis

ResearchDGX agent

arXiv:2608.20343v1 Announce Type: new Abstract: This study develops and evaluates a bankruptcy prediction framework that integrates consensus-based feature selection, hybrid resampling, stacking ensem

[Benchmark] Optimal DFlash2 quants for speed and context size, 5090 RTX, llama.cpp, Qwen 3.8 27B Dynamic3 Unsloth. Comparison with MTP

Model ReleasesDGX agent

Graph: A cumulative metric of Speed x Context Size - the higher the dot - the better. Helpful for looking for the most balanced solution. The research relies on DFlash2 Q2 work by AnalogAlok: https://

Beyond Gold Standards: Epistemic Ensemble of LLM Judges for Formal Mathematical Reasoning

ResearchDGX agent

arXiv:2506.10903v2 Announce Type: replace Abstract: Statement autoformalization plays a crucial role in formal mathematical reasoning by enabling the automatic translation of natural language statemen

Beyond Imitation: Self-Improving Robot Policies via Off-Policy Q-Planning

Model ReleasesDGX agent

arXiv:2608.21204v1 Announce Type: cross Abstract: Behaviour Cloning (BC) has driven remarkable progress in robot manipulation, yet it is fundamentally limited by its inability to self-improve: a polic

Beyond Raw Transcripts: Structured Persona Extraction for LLM-Based Digital Twins

Model ReleasesDGX agent

arXiv:2608.20344v1 Announce Type: new Abstract: LLM-based 'digital twins' aim to simulate how an individual would behavein new environments or respond to novel questions, given some representation of

Breaking High Confidence: Practical Face Impersonation under High-Security Thresholds

Model ReleasesDGX agent

arXiv:2608.20884v1 Announce Type: new Abstract: Face recognition systems (FRSs) are increasingly deployed in critical real-world services for authentication, such as banking applications and airport i

Coverage-Driven Verification for Safety-by-Design in AI-Based Collision Avoidance Systems

Model ReleasesDGX agent

arXiv:2608.20864v1 Announce Type: new Abstract: Artificial Intelligence (AI) offers significant potential for future aviation systems; however, its integration into safety-critical applications requir

Curriculum-Aware Interpolate-then-Refine: Learned Physiological Time-Series Imputation under Realistic Missingness

Model ReleasesDGX agent

arXiv:2608.21207v1 Announce Type: cross Abstract: Imputing physiological time series (arterial blood pressure, blood glucose, etc.) is essential for addressing the missingness that pervades clinical d

deepseek-v4-flash-0731 - surprisingly usable

Model ReleasesDGX agent

I just finished building my (relatively) low rent local inference machine: * Epyc 7663 * 256GB ECC DDR4-3200 * 1x RTX 5090 32GB Yeah I realize it's weird to throw a 5090 and 256GB of anything together

Doctor Rashomon and the UNIVERSE of Madness: Variable Importance with Unobserved Confounding and the Rashomon Effect

ResearchDGX agent

arXiv:2510.12734v2 Announce Type: replace Abstract: Variable importance (VI) methods are often used for hypothesis generation, feature selection, and scientific validation. In the standard VI pipeline

Empowering autonomous agents with advanced security governance

Model ReleasesDGX agent

AI agents are the ultimate insiders. We grant them permission to read emails, query databases, and trigger API calls. They don’t just retrieve information, they take action. Agents offer incredible po

Explainable Deepfake Detection with Feature-robust Augmentation and Evidence-grounded Explanation Optimization

ResearchDGX agent

arXiv:2608.20913v1 Announce Type: cross Abstract: Explainable deepfake detection extends binary classification by requiring models to not only predict authenticity but also provide interpretable justi

Explaining Intrinsic Moral Self-Correction with Mechanistic Interpretability

SafetyDGX agent

arXiv:2505.11924v4 Announce Type: replace-cross Abstract: Intrinsic moral self-correction refers to the phenomenon where a language model refines its ethical judgments or aligns its outputs purely thr

Federated and differentially private estimation of KL divergence

Model ReleasesDGX agent

arXiv:2411.16478v3 Announce Type: replace Abstract: Measuring distribution drifts is a key task in managing distributed, sensitive data, as it underpins a wide range of federated learning and analytic

Frozen CLIP Priors for Robust Self-Supervised Poisson Inverse Problems

Model ReleasesDGX agent

arXiv:2608.20524v1 Announce Type: cross Abstract: Self-supervised learning for imaging inverse problems is increasingly important in photon-limited settings, where acquiring clean ground truth is impr

Generalization Measures under Controlled Covariate Shift: A Regime-Aware Benchmark

Model ReleasesDGX agent

arXiv:2602.01718v2 Announce Type: replace Abstract: Predicting generalization from quantities available before target-test evaluation remains a central challenge in deep learning. The systematic bench

Intent Engine: Natural-Language Intent Translation for Intent-Driven Orchestration in the Compute Continuum

Model ReleasesDGX agent

arXiv:2608.20388v1 Announce Type: new Abstract: Microservice placement in the compute continuum is driven by low-level Service-level Objectives (SLOs), but requiring users to specify metric-level cons

Is Multimodal Speculative Decoding Ready for Diffusion-Based Parallel Drafting? A Survey and Empirical Diagnosis

SafetyDGX agent

arXiv:2608.20743v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by allowing a lightweight drafter to propose future tokens while a target model verifies them

Learning Prostate Anatomy at Test Time for Cancer Detection in Micro-Ultrasound

ResearchDGX agent

arXiv:2608.20557v1 Announce Type: new Abstract: Domain shift across clinical centers using different imaging hardware or acquisition protocols remains a fundamental barrier to deploying deep learning

On the Within-class Variation Issue in Alzheimer's Disease Detection

ResearchDGX agent

arXiv:2409.16322v4 Announce Type: replace-cross Abstract: Alzheimer's Disease (AD) detection commonly employs machine learning classification models to distinguish between individuals with AD and thos

Our 2nd founder dinner in SF co-hosted by @jerryjliu0 and @GuangyuRobert at @Fundamental - the team behind @tryshortcutai Talking about exis…

Model ReleasesDGX agent

Our 2nd founder dinner in SF co-hosted by @jerryjliu0 and @GuangyuRobert at @Fundamental - the team behind @tryshortcutai Talking about existing moats in the AI era. Frontier labs are moving past mode

PhotoBench: Beyond Visual Matching Towards Personalized Intent-Driven Photo Retrieval

Model ReleasesDGX agent

arXiv:2603.01493v2 Announce Type: replace-cross Abstract: Personal photo albums are not merely collections of static images but living, ecological archives defined by temporal continuity, social entan

Qwen 3.8 27B Aider score

Model ReleasesDGX agent

I ran the Aider benchmark on Qwen 3.8 27B FP8 with FP8 KV cache 256K context vLLM. The score: 72.9 This matches Gemini 2.5 Pro from 2025-04-12 which also scored 72.9. Beats Claude Opus 4 from 2025-05-

Qwen 3.8 27B, just wanted to say thanks to you guys

Model ReleasesDGX agent

I commented on another Qwen 3.8 27B post that I was frustrated getting anything to work. You all gave some great comments. I nuked openwebui and straightened out my llama.cpp docker config. 1 hour of

Qwen-3.8-27B, Nemotron-3.5-Lightning-30B-A3B, Ornith-1.5-35B-A3B, Muse-Glimmer-30B oQ8e comparison

Model ReleasesDGX agent

Ornith does really well. TielCoder (https://llm-bench.io/benchmarks/cmt7kp2zj002r01lcmpchvlko) might be even a bit better in coding. Will give it a try soon. Details of the comparison see here: https:

Real local agentic coding on a 12GB VRAM budget.

Model ReleasesDGX agent

Thanks to Unsloth Dynamic 3.0 quants coming in slightly leaner and better preserved, I settled on Qwen 3.8 27B (`UD_Q4_K_XL`) at 100K context as my daily driver for Hermes Agent and OpenCode. On an RT

RecGen3D: Reconstruction-Guided 3D Generation in a Shared Canonical Space

SafetyDGX agent

arXiv:2604.01479v3 Announce Type: replace Abstract: Sparse-view 3D modeling represents a fundamental tension between reconstruction fidelity and generative plausibility. While feed-forward reconstruct

Recognizing Artificial Minds: A Philosophical Defense of AI Cognition

ResearchDGX agent

arXiv:2504.13988v2 Announce Type: replace Abstract: This work defends the 'Whole Hog Thesis': sophisticated Large Language Models (LLMs) like ChatGPT are full-blown linguistic and cognitive agents, po

Research Paper Quality Recognition Through Textual Feature Analysis

Model ReleasesDGX agent

arXiv:2608.20368v1 Announce Type: new Abstract: Knowledge and innovations are shaped by using the quality and credibility of the scientific research. Yet, distinguishing between impactful, high-qualit

Roadside-Cooperative Autonomous Driving: From Data Platform to Vision-Language End-to-End Reasoning

Model ReleasesDGX agent

arXiv:2608.21032v1 Announce Type: new Abstract: Vehicle-to-Everything (V2X) cooperation enables beyond-line-of-sight perception, mitigating occlusions in single-vehicle sensing. However, existing V2X

Share the Judge, Learn the Deferral: Where Specialization Helps LLM Evaluation

AgentsDGX agent

arXiv:2607.27984v2 Announce Type: replace Abstract: Agentic systems generate outputs faster than human review. We contrast two LLM evaluator specialization strategies: specialized judge weights, or ru

Significant Other AI: Identity, Memory, and Emotional Regulation as Long-Term Relational Intelligence

ResearchDGX agent

arXiv:2512.00418v3 Announce Type: replace-cross Abstract: Significant Others (SOs) stabilize identity, regulate emotion, and support narrative meaning-making, yet many people today lack access to such

Sparse Token Routing in Efficient Transformers

Model ReleasesDGX agent

arXiv:2608.20632v1 Announce Type: new Abstract: Efficient-transformer research often motivates token pruning and adaptive computation with the claim that not all tokens require equal computational eff

The Cost of a Physics Prior Is Bounded by the Ablation Gap

ResearchDGX agent

arXiv:2608.21059v1 Announce Type: new Abstract: Shape-constrained and physics-informed learning reports an accuracy cost of enforcing a prior and treats it as a property of the prior. We show it is mo

The Exceedance Design Effect: Effective Sample Size for Thresholds under Clustering

Model ReleasesDGX agent

arXiv:2608.21262v1 Announce Type: cross Abstract: Many machine-learning systems set a threshold at a quantile of a calibration set: conformal predictors that promise 90% coverage by drawing their cuto

TreeWY: Speculative Verification for Gated DeltaNet Hybrids

ResearchDGX agent

arXiv:2608.20961v1 Announce Type: new Abstract: Modern open models are hybrids: most layers are linear-attention (Gated DeltaNet, GDN) layers carrying a small fixed-size recurrent state instead of a g

← Previous
1…501502503504505…1071
Next →