AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
Local Ai

HarnessWAM: Bridging Prediction and Deliberation in World Action Models

DGX agent

arXiv:2608.09516v1 Announce Type: new Abstract: World Action Models (WAMs) jointly learn environmental dynamics and robot actions, introducing priors over physical evolution into embodied control. How

local-aiarxiv-cs-ro
11 Aug 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

I gave DeepSeek V4 Flash basic vision by training a 40M connector on 100K examples

DGX agent

I wanted to find out whether a huge text-only MoE could be given basic vision without retraining the language model itself. The short answer is yes. I froze DeepSeek V4 Flash and a 417M-parameter Moon

model-releasesr-localllama
11 Aug 2026
Safety

Illusion of Alignment: Detecting Hidden Disagreement in Collaborative Dialogue

DGX agent

arXiv:2608.08210v1 Announce Type: new Abstract: Collaborative dialogue can end with apparent agreement while participants still differ on goals, assumptions, or execution plans, creating an extbf{illu

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models a…

DGX agent

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models are pushing the frontier of coding and knowledge work, but su

model-releasesjerry-liu--x
11 Aug 2026
Model Releases

LLM-Guided Heuristic Design from Simulation Traces: A Case Study in Dynamic Production and AGV Scheduling

DGX agent

arXiv:2608.09343v1 Announce Type: new Abstract: Simulation-based optimization (SBO) evaluates executable policies under stochastic dynamics, but most methods treat the simulator as a black box: aggreg

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

ML-Based Hierarchical Prediction for Practical Energy Scheduling in Dynamic NTN-WPT Systems

DGX agent

arXiv:2608.08804v1 Announce Type: cross Abstract: With advancements in long-distance wireless power transfer (WPT) and space-based energy technologies, integrating WPT into non-terrestrial networks (N

safetyarxiv-cs-lg
11 Aug 2026
Model Releases

Observations on Muse-Glimmer reasoning traces being noticeably different from qwen / gemma models and questions for you guys

DGX agent

Just downloaded the model, UD-Q5_K_XL quant, asked it to generate a long story to test out reasoning and speed with dflash (super fast btw, ~ 90 to 160 tok/s on a 5090 depending on task) and was surpr

model-releasesr-localllama
11 Aug 2026
Model Releases

Open source is so back. Zuck just announced Meta is opening the weights for Muse Glimmer, with Muse Spark 1.2 coming soon But a year ago, ev…

DGX agent

Open source is so back. Zuck just announced Meta is opening the weights for Muse Glimmer, with Muse Spark 1.2 coming soon But a year ago, everyone doubted Meta's position in the AI race In an intervie

model-releasesrowan-cheung--x
11 Aug 2026
Model Releases

OpenWALDO launches to build collaborative community for open-source AI

DGX agent

OpenWALDO, a new open-source artificial intelligence project sponsored by Ctrl IQ Inc., launched today, led by Gregory Kutzer, the founder of Rocky Linux, CentOS and Apptainer. The project aims to bui

model-releasessiliconangle
11 Aug 2026
Hardware

Personalized AI startup River AI raises $1.1B from consortium backed by Nvidia, AMD

DGX agent

River AI Inc., a startup that helps enterprises customize open-source artificial intelligence models, has raised 1.1 billion in early-stage funding. The company stated in today’s announcement that it

hardwaresiliconangle
11 Aug 2026
Safety

The Anatomy of a Prompt Injection: A Component Model for Structured Analysis

DGX agent

arXiv:2608.07808v1 Announce Type: cross Abstract: Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

💡The world needs an open-source platform, and that’s exactly what we’re building to give our customers more choice and the flexibility to c…

DGX agent

💡The world needs an open-source platform, and that’s exactly what we’re building to give our customers more choice and the flexibility to choose the right model for the right task. As part of this, we

model-releasesmistral-ai--x
11 Aug 2026
Model Releases

Whoa, Meta released a new open-weight LLM yesterday, something that hasn't happened since the good old Llama days. Their Meta Muse Glimmer m…

DGX agent

Whoa, Meta released a new open-weight LLM yesterday, something that hasn't happened since the good old Llama days. Their Meta Muse Glimmer model is a 30B multimodal reasoning model with a Gemma-like a

model-releasessebastian-raschka--x
11 Aug 2026
Model Releases

An Exploratory Evaluation of LLM-Assisted Rewriting of Moderate-Complexity Financial Sentences for DisCoCat-Based Sentiment Analysis

DGX agent

arXiv:2608.07439v1 Announce Type: new Abstract: Quantum natural language processing (QNLP) provides a grammar-aware framework for text modeling, and Distributional Compositional Categorical (DisCoCat)

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

Beyond Text Matching: Towards Reference-Free Evaluation for Human-Oriented Binary Reverse Engineering

DGX agent

arXiv:2608.07038v1 Announce Type: cross Abstract: Human-Oriented Binary Reverse Engineering (HOBRE) aims to transform decompiled pseudocode into a more human-friendly representation, thereby reducing

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Chat UIs with native audio input for multimodal models?

DGX agent

I've been running Gemma 4 E4B with oMLX and I can't find any chat interfaces that directly send the audio file to the model instead of running the audio through a separate STT layer. I can confirm the

model-releasesr-localllama
10 Aug 2026
Model Releases

CoBa: Cost-Effective Test-Time Scaling via Compute-Balanced Routing

DGX agent

arXiv:2608.07424v1 Announce Type: new Abstract: Test-time scaling is often implemented by spending more compute along one axis: sampling more solutions, extending a chain of thought, or applying a str

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Corma launches with $60M in funding for defensive cybersecurity AI

DGX agent

Defensive cybersecurity startup Corma Labs Ltd. today announced it has raised 60 million in seed funding to build a foundation model purpose-built for security defense. Founded in 2025, Corma runs off

model-releasessiliconangle
10 Aug 2026
Research

DocMemo: Dynamic Evidence Discovery via Probabilistic Memory-Guided Retrieval for Multi-Modal Document Understanding

DGX agent

arXiv:2608.07067v1 Announce Type: new Abstract: Long-document understanding requires locating sparse and heterogeneous evidence across hundreds of pages, yet existing systems remain limited by static

researcharxiv-cs-ai
10 Aug 2026
Model Releases

FinRank: An Evidence-Grounded Benchmark for Financial Question Answering and Retrieval over SEC Filings

DGX agent

arXiv:2608.07400v1 Announce Type: new Abstract: Financial question answering is typically evaluated by answer correctness, yet in SEC filings a plausible and even numerically correct answer can be gro

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks

DGX agent

arXiv:2608.07411v1 Announce Type: new Abstract: In the context of geodata, existing Large Language Models have often been studied in a homogeneous setting, which has considerably limited insights into

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Got Meta's new Muse Glimmer 30B running on my MacBook (M3 Max, 96GG) and tested the serving options available so far. Fastest right now: Oll…

DGX agent

Got Meta's new Muse Glimmer 30B running on my MacBook (M3 Max, 96GG) and tested the serving options available so far. Fastest right now: Ollama's MLX engine (DFlash included) at ~29 tok/s. Tuned llama

model-releasesollama--x
10 Aug 2026
Local Ai

How Malachyte solves retail’s cold-start problem with managed real-time AI

DGX agent

What’s the best way to recommend products to little-known users? We’ve spent our careers trying to solve this problem for major companies like Spotify and Priceline, and it’s why Sidd founded Malachyt

local-aigoogle-cloud-ai
10 Aug 2026
Safety

How WPP operationalizes platform and data engineering for AI marketing

DGX agent

Between chaotic levels of market fragmentation and economic volatility, marketing and communications agencies can no longer rely on the human intuition they’ve traditionally used to win clients and op

safetygoogle-cloud-ai
10 Aug 2026
Local Ai

Human-Centered Explainable AI for TinyML Edge Devices: A Pareto-Based Selection Framework with LLM-Guided Design

DGX agent

arXiv:2608.07091v1 Announce Type: cross Abstract: Edge Artificial Intelligence (Edge AI) enables the deployment of AI models directly on local edge devices, while such deployments are subject to stric

local-aiarxiv-cs-ai
10 Aug 2026
Model Releases

LitTraceQA: A Benchmark for Multi-Stage Grounding and Verification in Scientific Question Answering

DGX agent

arXiv:2608.07370v1 Announce Type: new Abstract: Scientific literature is increasingly used as a knowledge source for language models, retrieval-augmented generation systems, and research assistants, b

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers

DGX agent

arXiv:2608.06867v1 Announce Type: new Abstract: No single large language model (LLM) is optimal across all queries and budget constraints, making model routing essential for cost-effective deployment.

model-releasesarxiv-cs-cl
10 Aug 2026
Safety

MaskFlow: Precise, Consistent and Seamless Regional Image Editing

DGX agent

arXiv:2608.06929v1 Announce Type: cross Abstract: Regional image editing has attracted considerable attention for its spatial controllability. Although instruction-based and mask-reference-based editi

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

Meta released Muse Glimmer 30B: multimodal model for your Claw/Pi setups 🔥 we tested and fine-tuned the model for you, and shipped day-0 su…

DGX agent

Meta released Muse Glimmer 30B: multimodal model for your Claw/Pi setups 🔥 we tested and fine-tuned the model for you, and shipped day-0 support in transformers and llama.cpp, including DFlash for 2-4

model-releasesgeorgi-gerganov--x
10 Aug 2026
Model Releases

Meta releases open-source Muse Glimmer model with 30B parameters

DGX agent

Meta Platforms Inc. today released Muse Glimmer, an open-source language model that can run on personal computers. The company also published a lengthy essay penned by Chief Executive Officer Mark Zuc

model-releasessiliconangle
10 Aug 2026
Safety

MolBioKG: Grounding Out-of-Graph Molecules in Biomedical Knowledge Graphs via Multi-Resolution Structural Anchoring

DGX agent

arXiv:2608.06713v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) accelerate drug discovery, but standard pipelines assume query molecules already exist as graph entities, leaving unre

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

Muse Glimmer ACTUALLY fits on a single RTX 3090

DGX agent

I did some testing this morning, and I was surprised to find that Muse Glimmer actually comfortably fits on a single RTX 3090 with full context + DFlash + mmproj at Q4_K_XL, unlike Qwen3.6-27B and Gem

model-releasesr-localllama
10 Aug 2026
Local Ai

Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-G…

DGX agent

Muse Glimmer, the new 30B model, is available on Hugging Face right now - here's the GGUF version: https://huggingface.co/meta-models/Muse-Glimmer-30B-GGUF 1/ big announcement today: we will be releas

local-aisimon-willison--x
10 Aug 2026
Model Releases

Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accurate Optimization Modeling

DGX agent

arXiv:2608.07040v1 Announce Type: new Abstract: Solving combinatorial optimization problems (COPs) requires not only efficient algorithms but also carefully crafted formulations. While recent works ha

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Policy-Masked Private Experts: Auditable and Reversible Capability Access Control in Sparse MoE Models

DGX agent

arXiv:2608.06690v1 Announce Type: cross Abstract: Most language-model access controls regulate behavior while leaving the same computation available to every request. We study a different systems ques

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Social World Models

DGX agent

arXiv:2509.00559v3 Announce Type: replace Abstract: Humans intuitively navigate social interactions by simulating unspoken dynamics and reasoning about others' perspectives, even with limited informat

model-releasesarxiv-cs-ai
10 Aug 2026
Research

These startups are chasing the next big thing in LLMs

DGX agent

MIT Technology Review’s What’s Next series looks across industries, trends, and technologies to give you a first look at the future. You can read the rest of them here. Way back in the summer of 2017,

researchmit-tech-review
10 Aug 2026
Safety

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “…

DGX agent

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “who distributes it to the most people”: Instead of fighting

safetyyann-lecun--x
10 Aug 2026
Model Releases

Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML Comparisons

DGX agent

arXiv:2608.07303v1 Announce Type: new Abstract: Comparisons between AutoML systems at short time budgets -- tens of seconds rather than hours -- are common in tool READMEs and workshop papers, and the

model-releasesarxiv-cs-ai
10 Aug 2026
Local Ai

You can now try offline computer use with cua-driver + Muse Glimmer, through our friends at @ollama 🦙 How-to: https://cua.ai/docs/how-to-gu…

DGX agent

Francesco @francedot reported that the first fully offline computer‑based LLM experience was achieved by running Muse Glimmer 30B (from @AIatMeta) locally on macOS, using Cua Driver to control native

local-aiollama--x
10 Aug 2026
Model Releases

Zero Gap Is Not Restoration: Stratified Per-Question Probability Evaluation and Step-wise Mitigation of Benchmark Contamination

DGX agent

arXiv:2608.07341v1 Announce Type: cross Abstract: Test data from public benchmarks inevitably leaks into pretraining corpora, inflating evaluation scores once memorized. extbf{Contamination mitigation

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 Q8_K_XL sometimes stops mid-task in OpenCode - anyone else seeing this?

DGX agent

Hey everyone, I've been experimenting with the new DeepSeek-V4-Flash-0731 release locally using the Unsloth Studio Q8_K_XL GGUF with OpenCode. Overall, it's been working really well, but I've noticed

model-releasesr-localllama
9 Aug 2026
Industry

Forecasting the AI bubble: When scarcity turns to surplus

DGX agent

Artificial intelligence can be technologically transformative and still produce a capital bubble. Those two ideas are not in conflict. The bubble bursting does not require AI to fail. It only requires

industrysiliconangle
8 Aug 2026
Model Releases

Qwen 35B-A3B MoE vs 27B dense in local coding tests: ~4× faster, much smaller quality gap than I expected

DGX agent

I compared Qwen 35B-A3B MoE against Qwen 27B dense on a series of local coding-maintenance tasks. On my R9700/llama.cpp setup, the MoE model generated about 3.9× faster (~116 vs ~30 tok/s), but the co

model-releasesr-localllama
8 Aug 2026
Safety

Adaptive Arena-based Contestable Argumentative Network-of-Experts for Open-Ended Care Plan Coordination

DGX agent

arXiv:2608.05391v1 Announce Type: new Abstract: Care plan coordination demands synthesizing heterogeneous clinical, functional, and psychosocial information across multiple professional disciplines, w

safetyarxiv-cs-ai
7 Aug 2026
Tools

Autoscaling peaky LLM inference workloads is completely different than autoscaling something like a web service. I wrote a deepdive covering…

DGX agent

Zain (@zainhas) published a detailed article on August 7, 2026 explaining that autoscaling for highly peaky large‑language‑model (LLM) inference is fundamentally different from autoscaling conventiona

toolstogether-ai--x
7 Aug 2026
Model Releases

Codex is overoptimised for large models: it ranks 2nd out of 10 for GLM 5.2 but drops to 9th place for Gemma-4! Almost all the effort in thi…

DGX agent

Codex is overoptimised for large models: it ranks 2nd out of 10 for GLM 5.2 but drops to 9th place for Gemma-4! Almost all the effort in this field goes into tuning the weights. We wanted to know how

model-releasesclem-delangue--x
7 Aug 2026
Model Releases

Enhancing Social Intelligence in LLMs with Hierarchical Reasoning and Utterance-Level Goal Rewarding

DGX agent

arXiv:2608.05832v1 Announce Type: new Abstract: Large language models (LLMs) excel in structured tasks but struggle with dynamic social interactions, where success requires long-term goal coordination

model-releasesarxiv-cs-cl
7 Aug 2026
← Previous
1…301302303304305…371
Next →