AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlog
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,767 results
Tools

Real-time inference for robots at Physical Intelligence

DGX agent

Physical Intelligence (Pi) is building a general-purpose robotic intelligence system whose core Visual-Language-Action (VLA) model takes visual observations, natural-language instructions, and the...

toolsmodal-blog
8 Apr 2026
Tools

Safetensors is Joining the PyTorch Foundation

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

Safetensors, a secure tensor serialization format originally developed by Hugging Face, has joined the PyTorch Foundation as a foundation-hosted project under the Linux Foundation, alongside DeepSp...

toolshugging-face
8 Apr 2026
Model Releases

Strong release! GLM-5.1 is a DeepSeek-V3.2-like architecture (including MLA and DeepSeek Sparse Attention) but with more layers. And the ben…

DGX agent

Strong release! GLM-5.1 is a DeepSeek-V3.2-like architecture (including MLA and DeepSeek Sparse Attention) but with more layers. And the benchmarks look better throughout! Looks like THE flagship open

model-releasessebastian-raschka--x
8 Apr 2026
Model Releases

The feature I most want from AI labs right now is documentation on which underlying search engines they use when their chat tools run a sear…

DGX agent

The feature I most want from AI labs right now is documentation on which underlying search engines they use when their chat tools run a search OpenAI and Anthropic and Meta AI all have search and I ha

model-releasessimon-willison--x
8 Apr 2026
Model Releases

The new Anthropic managed agents API is basically the Letta API that we've had since a year ago, but closed source and with provider lock-in…

DGX agent

The new Anthropic managed agents API is basically the Letta API that we've had since a year ago, but closed source and with provider lock-in. They even have read-only memory blocks and memory block sh

model-releasesharrison-chase--x
8 Apr 2026
Model Releases

We partnered with @Zai_org to bring GLM-5.1 to Modal. Free to try as an endpoint for the next month. GLM-5.1 further improves upon GLM-5's c…

DGX agent

We partnered with @Zai_org to bring GLM-5.1 to Modal. Free to try as an endpoint for the next month. GLM-5.1 further improves upon GLM-5's coding abilities and long-horizon effectiveness. Introducing

model-releaseszhipu-ai--x
8 Apr 2026
Model Releases

got opencode working in taugentic, so now i can use glm-5.1 with my coding plan.

DGX agent

User Kevin Kern reported successfully integrating OpenCode into Taugentic (an AI-powered browser/agent environment), enabling use of GLM-5.1 via Z.ai's GLM Coding Plan. GLM-5.1 is Z.ai's next-gener...

model-releaseszhipu-ai--x
7 Apr 2026
Model Releases

How can you improve your agentic search pipeline? I just wrote a blog post with @tech_optimist from @lancedb to answer exactly that. TLDR: -…

DGX agent

How can you improve your agentic search pipeline? I just wrote a blog post with @tech_optimist from @lancedb to answer exactly that. TLDR: - Parse files and take page-level screenshots with LiteParse,

model-releasesjerry-liu--x
7 Apr 2026
Tools

I'm a big fan of the pelican GLM-5.1 drew me today, it even animated it! https://simonwillison.net/2026/Apr/7/glm-51/

DGX agent

Simon Willison tested Z.ai's GLM-5.1 using his standard 'pelican on a bicycle' SVG benchmark, and the model stood out by spontaneously generating a full HTML page with both the SVG and a separate ...

toolssimon-willison--x
7 Apr 2026
Model Releases

Join the ARC Prize team -- help us build ARC-AGI-4 and ARC-AGI-5

DGX agent

Join the ARC Prize team -- help us build ARC-AGI-4 and ARC-AGI-5 Platform Engineer - Benchmark Lead ARC Prize Foundation is hiring a senior engineer to build our benchmark platform * Expand ARC-AGI-3

model-releasesfrancois-chollet--x
7 Apr 2026
Applications

Oh no.

DGX agent

I was unable to retrieve the specific tweet at the URL provided (`https://x.com/emollick/status/2041600435320959330`). The tweet ID `2041600435320959330` does not appear in any search results, and ...

applicationsethan-mollick--x
7 Apr 2026
Industry

Technically all the vulnerabilities in public facing code are in the training data Make of that what you will

DGX agent

Technically all the vulnerabilities in public facing code are in the training data Make of that what you will Mythos Preview has already found thousands of high-severity vulnerabilities—including some

industryemad-mostaque--x
7 Apr 2026
Local Ai

v0.20.4-rc2: gemma4: Disable FA on older GPUs where it doesn't work (#15403)

DGX agent

Ollama v0.20.4-rc2 is a release candidate that addresses a compatibility issue with Flash Attention (FA) for the Gemma 4 model on older GPUs. CUDA versions older than 7.5 lack the support needed t...

local-aiollama-releases
7 Apr 2026
Model Releases

Visually rich documents are especially challenging for agents. Tables, charts, and images often break traditional document pipelines, making…

DGX agent

Visually rich documents are especially challenging for agents. Tables, charts, and images often break traditional document pipelines, making complex reasoning difficult📄 So we teamed up with @lancedb

model-releasesjerry-liu--x
7 Apr 2026
Model Releases

A Dual-Dimensional LLM Framework for Automated Item Incidental Content Similarity Analysis in Large-Scale Assessments

DGX agent

arXiv:2608.24825v1 Announce Type: new Abstract: The rapid expansion of large-scale assessments and the growing adoption of automatic item generation have intensified concerns about incidental content

model-releasesarxiv-cs-ai
26 Aug 2026
Model Releases

A Geometric Theory of Robust Fairness Audits

DGX agent

arXiv:2608.24818v1 Announce Type: new Abstract: Neighborhood-based fairness audits evaluate individual fairness by comparing predictions among similar individuals in feature space. Despite their wides

model-releasesarxiv-cs-lg
26 Aug 2026
Model Releases

Amortized Set Prediction for Inverse IFS Reconstruction from Density Maps

DGX agent

arXiv:2608.24175v1 Announce Type: new Abstract: Iterated Function Systems (IFS) generate self-similar fractals from a few contractive affine maps. The forward map from parameters to images is computat

model-releasesarxiv-cs-cv
26 Aug 2026
Model Releases

Anatomy of a Scam Call: What 10,000 real scam and spam calls reveal about how phone scammers operate

DGX agent

arXiv:2608.24127v1 Announce Type: cross Abstract: Telephone fraud is pervasive and costly, but its inner workings are rarely observed at scale. We analyze a complete corpus of 10,211 inbound scam and

model-releasesarxiv-cs-lg
26 Aug 2026
Safety

ASemConsist: Adaptive Semantic Feature Control for Training-Free Identity-Consistent Generation

DGX agent

arXiv:2512.23245v3 Announce Type: replace Abstract: Recent text-to-image diffusion models have significantly improved visual quality and text alignment. However, generating a sequence of images while

safetyarxiv-cs-cv
26 Aug 2026
Model Releases

At a 1M-token context length, QSA’s attention kernel is up to 7.6× faster in prefill and 4.9× faster in decode. With a 90% prefix-cache hit …

DGX agent

At a 1M-token context length, QSA’s attention kernel is up to 7.6× faster in prefill and 4.9× faster in decode. With a 90% prefix-cache hit rate, Qwen3.8-Flash-Next delivers 8.6× the prefill throughpu

model-releasesqwen--x
26 Aug 2026
Model Releases

Benchmarking LLM Judges for Voice-Agent Evaluation: Reliability, Calibration, and Human Oversight

DGX agent

arXiv:2608.24314v1 Announce Type: new Abstract: Evaluating conversational voice agents at scale re- quires reliable assessment methods that capture both observ- able interaction quality and the contex

model-releasesarxiv-cs-ai
26 Aug 2026
Model Releases

CAFE: Self-Improving Search Agents Need Co-Evolving Feedback

DGX agent

arXiv:2608.24794v1 Announce Type: new Abstract: Outcome-supervised search agents learn when and how to retrieve evidence, but terminal rewards neither localize intermediate errors nor redirect an ongo

model-releasesarxiv-cs-ai
26 Aug 2026
Applications

ConsensusTAS: Self-Supervised Temporal Action Segmentation for Long-Horizon Construction Videos

DGX agent

arXiv:2608.24043v1 Announce Type: new Abstract: Recognizing sequential construction activities is important for collaborative human-robot work; for example, robots are able to understand workers' curr

applicationsarxiv-cs-cv
26 Aug 2026
Local Ai

Contrastive Branch Policy Optimization

DGX agent

arXiv:2608.24300v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) enables language models to learn multi-turn interaction with external tools, yet its sparse outc

local-aiarxiv-cs-ai
26 Aug 2026
Local Ai

Do Recipes Have Personas? Characterizing and Generating Creator Style in Attributed Procedural Graphs

DGX agent

arXiv:2608.24369v1 Announce Type: new Abstract: While large language models (LLMs) possess vast zero-shot procedural knowledge, their tendency to produce homogenized logic often obscures the unique, i

local-aiarxiv-cs-ai
26 Aug 2026
Agents

EgoErrorVQA: Assess Egocentric Comprehension Capabilities through Procedural Errors for Ego-Agentic AI

DGX agent

arXiv:2608.24134v1 Announce Type: new Abstract: The majority of our everyday activities are procedural and consist of sequences of interdependent steps. However, existing benchmarks for Visual Agents

agentsarxiv-cs-cv
26 Aug 2026
Model Releases

Eluna: An Agentic LLM System for Automating Warehouse Operations with Reasoning and Task Execution

DGX agent

arXiv:2607.08960v2 Announce Type: replace-cross Abstract: Warehouse operations are governed by Standard Operating Procedures (SOPs) that encode complex, multi-system decision logic, which must be exec

model-releasesarxiv-cs-ai
26 Aug 2026
Research

Ethical LLM-Assisted Research: A Framework for Responsible Delegation, Verification, and Epistemic Value

DGX agent

arXiv:2608.23644v1 Announce Type: new Abstract: Large language models (LLMs) are becoming routine instruments of scientific research, assisting with literature synthesis, hypothesis development, codin

researcharxiv-cs-ai
26 Aug 2026
Model Releases

Fidelity Preference, Not Demographic Preference: A Pixel-Level Attribute-Sensitivity Audit of Image Aesthetic/Preference Scorers

DGX agent

arXiv:2608.23593v1 Announce Type: cross Abstract: Text-to-image systems use learned aesthetic scorers to filter training data and guide generation, but whether these scores encode demographic attribut

model-releasesarxiv-cs-ai
26 Aug 2026
Model Releases

GLM-5.3-Flash is live in Hermes Agent via Nous Portal https://portal.nousresearch.com

DGX agent

GLM-5.3-Flash is live in Hermes Agent via Nous Portal https://portal.nousresearch.com Introducing GLM-5.3-Flash - Leading capabilities at a highly competitive price - Natively multimodal with a 1M-tok

model-releasesnous-research--x
26 Aug 2026
Local Ai

GLM-5.3-Flash scores 57 on the Artificial Analysis Intelligence Index and is connected to GLM-5.3 along the Pareto frontier.

DGX agent

Zixuan Li reported that **GLM‑5.3‑Flash** achieved a score of **57 on the Artificial Analysis Intelligence Index**. The model is positioned on the Pareto frontier, indicating it extends its predecesso

local-aiollama--x
26 Aug 2026
Model Releases

Google’s new AI transcription edits out your ‘ums’ and ‘ahs’

DGX agent

Google has updated Gemini Audio with new transcription capabilities that automatically detect specialized jargon and more than 85 languages. Gemini 3.5 Transcribe is a new addition to the Gemini famil

model-releasesthe-verge-ai
26 Aug 2026
Safety

HAP: Head-Adaptive Visual Token Pruning via Cross-Modal Alignment

DGX agent

arXiv:2608.23921v1 Announce Type: new Abstract: Recent Vision-Language Models encode high-resolution images into long visual token sequences, incurring prohibitive prefill costs. To compress them, exi

safetyarxiv-cs-cv
26 Aug 2026
Model Releases

ICA: Information-Aware Credit Assignment for Visually Grounded Long-Horizon Information-Seeking Agents

DGX agent

arXiv:2602.10863v2 Announce Type: replace-cross Abstract: Long-horizon reinforcement learning for information seeking agents remains difficult because terminal rewards reveal whether the final answer

model-releasesarxiv-cs-ai
26 Aug 2026
Applications

Incorporating Cognitive Load and Knowledge Transfer for Multi-Domain Knowledge Tracing

DGX agent

arXiv:2608.24005v1 Announce Type: new Abstract: Knowledge Tracing (KT) aims to assess students' dynamic knowledge states from their learning histories. While most existing KT methods focus on single-d

applicationsarxiv-cs-ai
26 Aug 2026
Safety

Interpreting Control Latents for System Identification via Conditional Flow Matching

DGX agent

arXiv:2608.23887v1 Announce Type: new Abstract: Latent-conditioned adaptive policies can control robots across changing dynamics, but their learned latents remain internal representations of the polic

safetyarxiv-cs-ro
26 Aug 2026
Safety

Latent Dynamics-Aware OOD Monitoring for Trajectory Prediction with Provable Guarantees

DGX agent

arXiv:2603.14603v2 Announce Type: replace Abstract: In safety-critical Cyber-Physical Systems (CPS), trajectory prediction guides downstream planning and control. Deep learning models forecast well on

safetyarxiv-cs-ro
26 Aug 2026
Research

LumiXAI: A Modular Full-Stack Framework for Feature Attribution

DGX agent

arXiv:2608.24524v1 Announce Type: cross Abstract: Feature attribution is a central tool of model interpretability, yet the software through which it is applied remains fragmented: individual tools spe

researcharxiv-cs-ai
26 Aug 2026
Research

Memory Is Not Always Needed: Characterizing Conditional Memory in Scientific Reasoning

DGX agent

arXiv:2608.23982v1 Announce Type: new Abstract: Scientific reasoning requires language models to retrieve specialized knowledge and incorporate it reliably into multi-step computation. Conditional mem

researcharxiv-cs-ai
26 Aug 2026
Safety

MetaRAG: Belief-Action Aligned Policy Optimization for Agentic RAG

DGX agent

arXiv:2608.24214v1 Announce Type: new Abstract: Agentic retrieval-augmented generation (RAG) requires language models to decide when to continue searching and when to answer. Existing RL-based methods

safetyarxiv-cs-ai
26 Aug 2026
Model Releases

MoE-based Feature Adapter for Prompt-free Binary Coronary Artery Segmentation in X-ray Angiography

DGX agent

arXiv:2608.24783v1 Announce Type: new Abstract: Accurate segmentation of coronary arteries in X-ray angiography videos is essential for quantitative coronary analysis and image-guided interventions. H

model-releasesarxiv-cs-cv
26 Aug 2026
Model Releases

MoRF-AST: Calibrated Probabilistic Virtual Sensing for Structural Monitoring under Changing Operating Conditions

DGX agent

arXiv:2608.24531v1 Announce Type: cross Abstract: Probabilistic full-field reconstruction provides uncertainty-aware response evidence for structural reliability assessment, yet inference from sparse

model-releasesarxiv-cs-lg
26 Aug 2026
Safety

NeuronGuard: Robust LLM Safety Alignment via Ablation-Aware Safety Signal Redistribution

DGX agent

arXiv:2608.23959v1 Announce Type: cross Abstract: Safety alignment in large language models (LLMs) remains brittle against a growing spectrum of attacks. Jailbreak attacks bypass safety mechanisms thr

safetyarxiv-cs-ai
26 Aug 2026
Model Releases

No reward hacking was found in GLM-5.2: solve the task, not the benchmark.

DGX agent

No reward hacking was found in GLM-5.2: solve the task, not the benchmark. Introducing reward hacking score corrections to the Artificial Analysis Coding Agent Index In v1.4 of the Artificial Analysis

model-releasesollama--x
26 Aug 2026
Model Releases

Paritok-4B: Intent-Conditioned Context Compression for Coding Agents

DGX agent

arXiv:2608.24188v1 Announce Type: new Abstract: Coding agents re-send large file reads and tool outputs to a frontier LLM every turn, and this context dominates their token bill. General-purpose promp

model-releasesarxiv-cs-ai
26 Aug 2026
Research

PROOF-Gen: From Optimized Data to Better Distillation

DGX agent

Supervised fine-tuning on teacher-generated trajectories is the standard first stage for distilling tool-calling capabilities into deployable models. Post-training pipelines that drive shipped tool-ca

researchapple-ml-research
26 Aug 2026
Model Releases

QABBA: Error-Guaranteed Symbolic Time-Series Compression via Integer-Quantized Aggregation

DGX agent

arXiv:2411.15209v3 Announce Type: replace Abstract: The expansion of time-series data from sensors and monitoring systems has made compact representations increasingly important. Such representations

model-releasesarxiv-cs-lg
26 Aug 2026
Model Releases

SA-Bench: Evaluating Semantic Alignment in LLM-Based Paper Reproduction

DGX agent

arXiv:2608.24252v1 Announce Type: new Abstract: LLM agents can generate paper reproduction code, yet often produce scientifically unfaithful implementations. We define this failure mode as semantic dr

model-releasesarxiv-cs-ai
26 Aug 2026
← Previous
1…642643644645646…1371
Next →