AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,648 results
27 Jul 2026

Token-Operations-Oriented Inference Optimization Techniques for Large Models

ApplicationsDGX agent

arXiv:2606.20295v2 Announce Type: replace-cross Abstract: Large model inference optimization serves as a key foundation for supporting the scalable, low-cost, and highly stable operation of large mode

Toward Goal-Agnostic Joint-Embedding Predictive Control of Partial Differential Equations

Model ReleasesDGX agent

arXiv:2607.21644v1 Announce Type: new Abstract: We present a goal-agnostic control framework for partial differential equations (PDEs) built around a joint-embedding predictive architecture (JEPA). Th

Toward High-Fidelity 3D Point-Cloud Learning for Brain Folding Morphology Prediction Using Trans-Unet

Local AiDGX agent

arXiv:2607.21840v1 Announce Type: new Abstract: Learning high-fidelity point-cloud features in the 3D space poses significant challenges, including permutation invariance, lack of local context, diffi

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions

Model ReleasesDGX agent

arXiv:2607.21635v1 Announce Type: new Abstract: Personal agents maintain memories, learned skills, tool configurations, and policy state that evolve with each user. Existing agent benchmarks often eva

Towards Reducing Foreign Language Anxiety Using Level-Appropriate Embodied Conversational Agents

AgentsDGX agent

arXiv:2607.21887v1 Announce Type: cross Abstract: Foreign language anxiety (FLA) can be a major barrier to second language acquisition (SLA), especially in conversational contexts. With the proliferat

TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI

SafetyDGX agent

arXiv:2607.22465v1 Announce Type: cross Abstract: Routing to select large language models (LLMs) with different cost-quality trade-offs has become a fundamental deployment feature of enterprise AI. Ex

Trajectory-Regularized Stochastic Optimal Control via KL Divergence

Model ReleasesDGX agent

arXiv:2607.22201v1 Announce Type: cross Abstract: We introduce trajectory-regularized stochastic optimal control (TRSOC), which augments standard stochastic optimal control (SOC) with a Kullback--Leib

TRaM-VSR: Importance-Aware Token Routing and Merging for One-Step Diffusion Video Super-Resolution

Local AiDGX agent

arXiv:2607.22231v1 Announce Type: new Abstract: Video super-resolution (VSR) using large-scale Diffusion Transformer (DiT) priors achieves exceptional perceptual quality but is often impractical due t

TriGlue: a Biology-Inspired Generative Model for Generating Molecular Glue-Induced Ternary Complex

ResearchDGX agent

arXiv:2607.22143v1 Announce Type: new Abstract: Molecular glue degraders have emerged as a promising strategy for targeted protein degradation by inducing ternary complex formation between an E3 ubiqu

Trying out LoKr instead of LoRA on Krea2

Local AiDGX agent

Dataset of 43 images, captioned with qwen3 VL 4B instruct, 50 word caption focusing on: Composition, Subject's hair, expression, clothes, pose, background Training Parameters: (10 rep x 43 image) x 6

Twins: Learn to Predict Unified Representations with Focal Loss

SafetyDGX agent

arXiv:2607.22531v1 Announce Type: new Abstract: Unified multimodal models seek a shared visual token space that supports both multimodal understanding and image generation. Discrete methods unify the

Unable to get GPU Passthrough working - Docker

Local AiDGX agent

Setup as follows: Proxmox -> Debian -> Docker -> Ollama. Other containers work. Compose file contains gpu device. Does it need nvidia runtime or any other options? If someone could provide an example

Unbiased Open World Regularization for Fair Self-Supervised Learning

Model ReleasesDGX agent

arXiv:2607.22149v1 Announce Type: new Abstract: Despite recent advances, self-supervised learning (SSL) models and Joint-Embedding Predictive Architectures (JEPAs) remain susceptible to learning spuri

Unboxing Diffusion Models for the Arts: Interactive Model Bending and Practice-Based Explainability

ResearchDGX agent

arXiv:2607.22428v1 Announce Type: cross Abstract: Explainable AI (XAI) in creative practice can be less about technocentric explanation and more about enabling artists to inspect modify and debug mode

Unexpected use of local llm

Local AiDGX agent

I was refreshing my youtube and found out my favourite reviewer uploaded a battery test of 78 smartphones: https://youtu.be/MpgUFrsIWSQ the author said they started using robotic arm to simulate a per

Unified Static-Dynamic Pruning for Efficient LLM Inference

HardwareDGX agent

arXiv:2607.21985v1 Announce Type: cross Abstract: The increasing deployment of large language models (LLMs) has magnified the computational and memory bottlenecks of autoregressive decoding, where low

Universal BCI Personalization: One API for Frozen EEG Trunks and Foundation Models

ResearchDGX agent

arXiv:2607.22397v1 Announce Type: cross Abstract: Frozen EEG encoders proliferate; per-model fine-tune defaults do not scale. We present Nimbus Personalizer: one contract encode to Bayesian head to Br

US tech giants have shifted their stance to publicly backing open AI models, with Anthropic and Amazon remaining notable holdouts alongside the US government (M.G. Siegler/Spyglass)

IndustryDGX agent

M.G. Siegler / Spyglass: US tech giants have shifted their stance to publicly backing open AI models, with Anthropic and Amazon remaining notable holdouts alongside the US government — The open letter

v0.32.5

Local AiDGX agent

**Ollama – v0.32.5 Release Summary** - Version **v0.32.5** (released 27 Jul at 01:25) is the latest stable release on GitHub, with a signed commit (GPG Key ID B5690EEEBB952194). - The update includes

Variance-Reduced Q-Learning over Static and Time-Varying Networks

TutorialsDGX agent

arXiv:2607.21876v1 Announce Type: new Abstract: We investigate a decentralized reinforcement learning problem involving multiple agents that interact with the same Markov Decision Process (MDP). The a

Variational Low-rank Tensor Decomposition for Multisubject Spatiotemporal Data Analysis

Model ReleasesDGX agent

arXiv:2607.22262v1 Announce Type: cross Abstract: Modeling shared and subject-specific structure in multisubject spatiotemporal data remains challenging, particularly in neuroimaging, where both spati

Vector-Valued Reproducing Kernel Banach Spaces for Neural Networks and Operators

ResearchDGX agent

arXiv:2509.26371v3 Announce Type: replace-cross Abstract: Recently, there has been growing interest in characterizing the function spaces underlying neural networks. While shallow and deep scalar-valu

Very cool paper from Microsoft. The idea is to train agents on replayed teacher trajectories instead of live environment rollouts. On-policy…

SafetyDGX agent

Very cool paper from Microsoft. The idea is to train agents on replayed teacher trajectories instead of live environment rollouts. On-policy distillation for agentic tasks is expensive because every u

Visual Relocalization from Sparse Views in Aliased and Low-Texture Environments via Novel View Synthesis

Local AiDGX agent

arXiv:2607.22147v1 Announce Type: new Abstract: Visual localization becomes extremely challenging in planetary-like terrains characterized by low texture, perceptual aliasing, harsh illumination, and

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning

TutorialsDGX agent

arXiv:2607.22013v1 Announce Type: new Abstract: Multimodal chain-of-thought (CoT) reasoning integrates visual and textual cues through step-by-step inference. In small models with limited token budget

ViTacWorld: Scaling Visuo-Tactile World Models for Contact-Rich Robot Manipulation

SafetyDGX agent

arXiv:2607.22530v1 Announce Type: new Abstract: Contact-rich robot manipulation requires physical interaction cues that are often invisible to cameras, making tactile sensing essential for robust cont

VTM-Nav: Harnessing Cross-Episode Experience for Object-Goal Navigation with Hierarchical Visual-Topological Memory

AgentsDGX agent

arXiv:2607.14514v2 Announce Type: replace Abstract: Training-free ObjectNav agents increasingly use vision-language models (VLMs), yet typically discard acquired scene knowledge after each request. We

Want to go deeper? Join Moonshot AI and Together AI for a technical webinar on how K3 was built and how to use it for production agent workf…

AgentsDGX agent

Together AI has released the Kimi K3 model on its platform as a Day‑0 launch partner for Moonshot AI’s open frontier agentic model, which supports long‑running workflows across code, tools, vision and

Wasserstein Gradient Flows for Scalable and Regularized Barycenter Computation

ResearchDGX agent

arXiv:2510.04602v4 Announce Type: replace-cross Abstract: Wasserstein barycenters provide a principled approach for aggregating probability measures, while preserving the geometry of their ambient spa

Way Security, which uses AI-driven automation and agentic workflows to help deploy IAM systems, raised a $20M seed from Insight Partners and Glilot Capital (Chris Metinko/Axios)

AgentsDGX agent

Chris Metinko / Axios: Way Security, which uses AI-driven automation and agentic workflows to help deploy IAM systems, raised a 20M seed from Insight Partners and Glilot Capital — Way Security raised

We are in a world where you can create truly unique, visually interesting and creative playable demos on demand with the current capabilitie…

Model ReleasesDGX agent

We are in a world where you can create truly unique, visually interesting and creative playable demos on demand with the current capabilities of Codex and Claude Code. We don't need to keep cloning th

We are releasing PerceptionBench, a benchmark that isolates visual perception and evaluates it as a set of atomic capabilities - discovered …

Model ReleasesDGX agent

We are releasing PerceptionBench, a benchmark that isolates visual perception and evaluates it as a set of atomic capabilities - discovered from how today's models fail, rather than defined in advance

We could really use Qwen3.8 in 27B, 35B, 122B and 397B sizes

Model ReleasesDGX agent

Instead of 2T+ models, continuing to release highly capable small to medium size LLMs would really help to keep this community vibrant. Hardly anyone can even dream of running the recent 1.5-2T+ beast

We're proud to join @NVIDIA and the Open Secure AI Alliance. To support open source models, we're contributing our research on measuring the…

HardwareDGX agent

We're proud to join @NVIDIA and the Open Secure AI Alliance. To support open source models, we're contributing our research on measuring the trustworthiness and security of open source models. Closing

We've joined the alliance. Open-weight models will ensure that we live in a safer digital world, and that America does not get left behind

IndustryDGX agent

We've joined the alliance. Open-weight models will ensure that we live in a safer digital world, and that America does not get left behind Attackers have frontier AI. Defenders need a frontier AI ecos

We've open-sourced AgentENV in collaboration with kvcache-ai. AgentENV is a distributed system for running agent environments at scale. Its …

AgentsDGX agent

We've open-sourced AgentENV in collaboration with kvcache-ai. AgentENV is a distributed system for running agent environments at scale. Its components power agentic RL training for Kimi K3, with fast

We've open-sourced MoonEP, our high-performance communication library for distributed MoE workloads. Built to make expert-parallel communica…

Model ReleasesDGX agent

We've open-sourced MoonEP, our high-performance communication library for distributed MoE workloads. Built to make expert-parallel communication more efficient at scale, MoonEP helps reduce communicat

What Happens to Accuracy When Photo Lineups Contain Non-Mated Rank-One Images From Large Galleries?

ResearchDGX agent

arXiv:2607.21792v1 Announce Type: new Abstract: One-to-many facial identification is commonly used to match a probe image from surveillance video against a gallery of driver's licenses and/or booking

What Matters When Building Universal Multilingual Named Entity Recognition Models?

ResearchDGX agent

arXiv:2601.06347v2 Announce Type: replace Abstract: Recent progress in universal multilingual named entity recognition (NER) has been driven by multilingual transformer models, task-specific architect

WHBench: Evaluating Frontier LLMs with Expert-in-the-Loop Validation on Women's Health Topics

Model ReleasesDGX agent

arXiv:2604.00024v2 Announce Type: replace Abstract: Large language models are increasingly used for medical guidance, but women's health remains under-evaluated in benchmark design. We present the Wom

When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas

SafetyDGX agent

arXiv:2505.19212v2 Announce Type: replace Abstract: Recent advances in LLMs have enabled their use in complex agentic roles, involving decision-making with humans or other agents, making ethical align

Why Large Language Models and Humans Converge and Diverge in Evaluating Creativity

SafetyDGX agent

arXiv:2607.22218v1 Announce Type: new Abstract: Despite the growing use of large language models (LLMs) as creativity evaluators, evidence of their alignment with human evaluations remains mixed, rais

With Comfy MCP, workflows can be built from anywhere. Describe the idea, let an AI agent build it, and check back when it's done. To try thi…

AgentsDGX agent

Comfy MCP enables users to create ComfyUI workflows from any location by simply describing their ideas. An AI agent will automatically build the workflow, after which the user can monitor progress unt

Would I be too cynical in thinking that the whole thing might fall apart if Nvidia wasn’t subsidizing? 🤔

HardwareDGX agent

Would I be too cynical in thinking that the whole thing might fall apart if Nvidia wasn’t subsidizing? 🤔 JUST IN : NVIDIA IN TALKS TO PROVIDE 250 BILLION FINANCIAL BACKSTOP FOR OPENAI DATA CENTER IN O

You can now fine-tune my 3.96M-parameter TTS on your own voice or language

Model ReleasesDGX agent

When I released Inflect v2 last week, I thought most people would ask whether a TTS model this small actually sounded decent. Instead, I kept getting two questions: “Can I train it on my own voice?” “

Yugabyte targets the missing memory and knowledge layer for enterprise AI agents

AgentsDGX agent

Enterprise investment in agentic artificial intelligence is accelerating, but the infrastructure supporting those systems is still catching up. Organizations are moving agents into customer support, s

Zero-Shot Mission-Level Evaluation for Aerial MLLM Agents

Model ReleasesDGX agent

arXiv:2607.22014v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are emerging as core reasoning modules for embodied agents, yet it remains unclear how well general-purpose m

26 Jul 2026

100B Models on Cheap Hardware: how realistic and limitations

Model ReleasesDGX agent

There is a lot of buzz around running 100B parameter models on cheap local hardware using ternary (1.58-bit) quantization like Microsoft's BitNet architecture. The theoretical hardware shortcuts are i

16 bit better than lower quants for Qwen3.6-27B

Model ReleasesDGX agent

I am writing a fairly complex C++ windows MFC application. I have a few 3090s and can run F16 Qwen3.6-27B with 256K context and MTP. The quality of code is exceptional with this quant vs its lower qua

23 Gemma4-E4B models compared with abliterlitics: the most downloaded one is also the most broken

Model ReleasesDGX agent

This is our biggest comparison yet. We've taken 23 Gemma 4 E4B models from huggingface and ran them through the abliterlitics gauntlet. We also have a new abliterlitics discord, feel free to jump on a

90 agentic bakeoff runs: ThinkingCap vs Fable Fusion vs stock Qwen3.6-27B

Model ReleasesDGX agent

Last week someone here said ThinkingCap and Fable Fusion 'really do beat the OG' for agentic work, so I ran it: 6 self-grading tasks, 5 reps, 3 models, 90 isolated runs. Tooling, since that's half the

again, the log is the agent

AgentsDGX agent

The thread argues that applied AI has entered a distributed‑systems phase, with event‑driven architectures treating logs as “agents” that consume data rather than serve as endpoints. It notes that alm

ai-sage/GigaChat3.1-Audio-10B-A1.8B · Hugging Face

Local AiDGX agent

GigaChat Audio 10B is an audio-native LLM built on top of the GigaChat 3.1 Lightning text model. A Conformer speech encoder and a modality adapter feed audio embeddings directly into a Mixture-of-Expe

An Inside Look at the Relay Market Powering Token Resellers and Fraud

ToolsDGX agent

An Inside Look at the Relay Market Powering Token Resellers and Fraud Fascinating investigation by Matt Lenhard into the market that has grown up around reselling LLM tokens at a discount by pooling A

Announcing the Claude Code-compatible interface for our new Fugu-Ultra v1.1! 🐡 Put a dynamically coordinated team of frontier models to wor…

Model ReleasesDGX agent

Announcing the Claude Code-compatible interface for our new Fugu-Ultra v1.1! 🐡 Put a dynamically coordinated team of frontier models to work inside the coding workflow you already know. Instead of rel

b10141

Model ReleasesDGX agent

mtmd: fix android build (#26150) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubunt

BeeLlama.cpp v0.4.1: KVarN, KV precision tail, q2_0-q3_1 KV cache, improved support. KLD benchmarks: tail 1024 makes kvarn5 and q6_0 match q8_0, for much less VRAM

Model ReleasesDGX agent

TL;DR llama.cpp fork with more KV cache quantization features, with all claims supported by benchmarks: KVarN, KV cache precision tail, additional types of standard KV cache (q2_0-q3_1, q6_0, q6_1), a

BREAKING: A Redditor just discovered that shared Claude conversations have been showing up in public search results. The post has 4K upvotes…

Model ReleasesDGX agent

BREAKING: A Redditor just discovered that shared Claude conversations have been showing up in public search results. The post has 4K upvotes and hundreds of comments, so this is spreading fast. Here's

CEO of Hugging Face: 'In the spirit of transparency, here’s what I asked OpenAI'

Local AiDGX agent

clem 🤗 on 𝕏: https://x.com/ClementDelangue/status/2081056675558195657 • Radical transparency: let’s release the traces from the “rogue” agents so the entire research community can study what happened.

Do people building local LLM rigs track RTX Ada/workstation card prices, or just consumer cards like the 5090?

Local AiDGX agent

curious how people here approach buying high-end/workstation cards (RTX 6000 Ada, 5000 Ada, etc) for local LLM work, do you actively watch pricing/timing on these specifically, or is the consumer 5090

← Previous
1…201202203204205…1411
Next →