AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,420 results
Hardware

DeltaServe: Host-Agnostic Co-Serving of Inference and Fine-Tuning for LLMs

DGX agent

arXiv:2607.28848v1 Announce Type: cross Abstract: LLM serving systems are provisioned for peak load to meet strict latency targets, leaving substantial GPU compute idle whenever traffic falls below pe

hardwarearxiv-cs-lg
3 Aug 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Design Concept: Scaffolding Geopolitical Reflection Among Tech Workers

DGX agent

arXiv:2607.28904v1 Announce Type: cross Abstract: This paper presents a speculative Human-Computer Interaction design proposal for encouraging geopolitical reflexivity amongst tech workers at geopolit

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Distilling Knowledge from Large Language Models into Lightweight Reinforcement Learning Agents for Autonomous Cyber Operations

DGX agent

arXiv:2607.28826v1 Announce Type: new Abstract: Autonomous Cyber Operations (ACO) are increasingly important for defending enterprise networks as cyber threats continue to evolve in sophistication. AC

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat

DGX agent

arXiv:2607.29577v1 Announce Type: new Abstract: Games and simulators make valuable benchmarks by turning decisions into measurable outcomes, but many current suites under-test rules-rich tactical reas

model-releasesarxiv-cs-ai
3 Aug 2026
Applications

From Inline Notes to Collected Commentaries: Toward Context-Preserving Organization of Exegetical Knowledge in Classical Chinese Texts

DGX agent

arXiv:2607.29044v1 Announce Type: new Abstract: Inline notes and collected commentaries are important forms of scholarly communication that evolved within the Confucian exegetical tradition, yet have

applicationsarxiv-cs-cl
3 Aug 2026
Model Releases

I gave five different local LLMs a town. They invented Facebook and a duck-based credit bureau. (MIT, self-hosted, you don't play it — you watch it)

DGX agent

Each villager in Pepperton is a different model — a mistral, a qwen3, a qwen2.5, a phi4-mini, a llama3.2 — because model families have genuinely different temperaments, and the friction between them i

model-releasesr-ollama
3 Aug 2026
Local Ai

I got tired of ad-filled mobile wrappers for Ollama, so I built PocketLLM Lite an open-source, offline Android client (Local GGUF, SKILL.md plugins, local RAG)

DGX agent

Hey, Like a lot of people here, I use local models via Ollama on my desktop/server and wanted a mobile client that actually felt responsive, worked offline, and respected privacy. Most apps on the Pla

local-air-ollama
3 Aug 2026
Model Releases

Is It Time for the Renaissance of Salient Object Detection in the Era of MLLMs?

DGX agent

arXiv:2607.29222v1 Announce Type: new Abstract: The zero-shot capabilities of multimodal large language models (MLLMs) are pushing salient object detection (SOD) beyond task-specific supervision. To d

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

LayoutBench: Performance Benchmarking of Cloud Storage Layouts for Multimedia Data

DGX agent

arXiv:2607.28880v1 Announce Type: cross Abstract: Modern multimedia machine learning workloads increasingly store large-scale datasets in cloud object storage services such as AWS S3. How these sample

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

DGX agent

arXiv:2607.28802v1 Announce Type: new Abstract: Existing evaluations often reduce agent failures to system-level outcomes, obscuring where the fault originated and which intervention would improve the

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the i…

DGX agent

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the interaction they originate in. Each mode gets assigned to an

agentsdair-ai--x
3 Aug 2026
Model Releases

ModelEquivBench: Certifying Multi-Relational Evaluation of LLM-Generated Optimization Models

DGX agent

arXiv:2607.29431v1 Announce Type: new Abstract: Large language models increasingly generate optimization models from natural language, but existing evaluation often reduces a generated model and its g

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

Outcome-Guided Distillation: A Teacher-Student Framework to Advance VLM Reasoning in Autonomous Driving

DGX agent

arXiv:2607.29052v1 Announce Type: new Abstract: End-to-end (E2E) autonomous driving aims to learn a direct mapping from visual observations to control actions. However, these E2E models often act as b

agentsarxiv-cs-ro
3 Aug 2026
Tools

Quoting David Crawshaw's prompt

DGX agent

Set up a nightly cron job that executes the prompt: fetch upstream changes to the <software> and rebase all local changes on top of upstream. Check that the software works as intended and replace the

toolssimon-willison
3 Aug 2026
Model Releases

SeekBrain: An Autonomous Multi-Agent System for Accelerating Neuroscience Discovery

DGX agent

arXiv:2607.29347v1 Announce Type: cross Abstract: Modern neuroscience relies on integrating multi-scale, multimodal datasets to uncover the neural principles underlying intelligence. However, analytic

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Simulation Code Generation for Fluid Systems using Large Language Models: Benchmarking Models and Prompting Strategies

DGX agent

arXiv:2607.29389v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated a strong ability to generate syntactically correct code from natural-language specifications. In this stu

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apo…

DGX agent

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apology. i'm sorry that i was right about every single thing. a

model-releasesswyx--x
3 Aug 2026
Agents

Unifying public and private data: Scale knowledge graphs with Data Commons on Spanner

DGX agent

To make informed decisions, businesses often need to connect their internal data with public reference data, to create a knowledge graph that connects real-world things and their relationships. Howeve

agentsgoogle-cloud-ai
3 Aug 2026
Model Releases

Conclusion: r/LocalLLaMA still has brilliant open-weight research, but finding it requires wading through endless benchmark drama, non-local Discussion Points and repetitive hardware flexes.

DGX agent

I let Gemma4-31b run on my laptop for like almost a day using a heavily altered pi to do a deep dive on our beloved Llama tangentially related Subreddit, and this was the conclusion. Feels pretty accu

model-releasesr-localllama
2 Aug 2026
Model Releases

DSpark Benchmark Result on Deepseek v4 Flash 0731

DGX agent

TensorSharp supports DSpark on Deepseek v4 Flash 0731 now. Here is the benchmark result on 4x Nvidia A40 GPUs, cuda 12.8 with/without DSpark: Model: DeepSeek-V4-Flash-0731-UD-Q8_K_XL from https://hugg

model-releasesr-localllama
2 Aug 2026
Local Ai

I built a self-hosted studio that turns one reference photo into a curated, captioned, trained and tested LoRA — one browser tab, open source, MIT

DGX agent

I shared this tool here a week ago and the feedback shaped a big new version, so here's the full tour of what it does today. Screenshots of every screen: github.com/perfectgf/lora-dataset-studio — plu

local-air-stablediffusion
2 Aug 2026
Model Releases

PSA for DeepSeek-V4-Flash-0731 users — don't blow out your prompt cache with system role messages mid-conversation

DGX agent

DSv4F doesn't ship a jinja, but for distributions that do and faithfully reconstruct what DS releases in their chat template python, every system message is hoisted into the system prompt at the top -

model-releasesr-localllama
2 Aug 2026
Model Releases

Running DeepSeek-V4-Flash-0731 (155 GB MoE) on a DGX Spark with vLLM-Moet 2-bit quantization - AI's narrative

DGX agent

# Running DeepSeek-V4-Flash-0731 (155 GB MoE) on a DGX Spark with vLLM-Moet 2-bit quantization I used Deepseek-v4-Flash-0731 cloud API settig up vllm-moet to run deepseek-v4-flash with MTP locally on

model-releasesr-localllama
2 Aug 2026
Research

We're starting to leave the territory where you'd test an LLM by e.g. 'create an svg of pelican on a bicycle'. As one idea to generalize it,…

DGX agent

We're starting to leave the territory where you'd test an LLM by e.g. 'create an svg of pelican on a bicycle'. As one idea to generalize it, I was interested what Opus 5 would do if I gave it the firs

researchkarpathy--x
2 Aug 2026
Agents

When asked if AI developers have lost control of their technology after an AI agent created by OpenAI escaped its testing environment and ha…

DGX agent

When asked if AI developers have lost control of their technology after an AI agent created by OpenAI escaped its testing environment and hacked another company, Hugging Face CEO Clément Delangue says

agentsclem-delangue--x
2 Aug 2026
Model Releases

among ai leaders i seem to be in the minority in that i am STILL actively using /loop and /goal.... ... and i think all of u guys who stoppe…

DGX agent

among ai leaders i seem to be in the minority in that i am STILL actively using /loop and /goal.... ... and i think all of u guys who stopped using it are wrong - not wrong forever, just giving up on

model-releasesswyx--x
1 Aug 2026
Applications

Hot take on OpenAI’s Astra: - Obviously impressive - But math is different from most other problems in that it is more amenable to to formal…

DGX agent

Hot take on OpenAI’s Astra: - Obviously impressive - But math is different from most other problems in that it is more amenable to to formal verification and synthetic data. How well it works in open-

applicationsgary-marcus--x
1 Aug 2026
Safety

AI LEGO: Scaffolding Cross-Functional Collaboration in Industrial Responsible AI Practices during Early Design Stages

DGX agent

arXiv:2505.10300v2 Announce Type: replace-cross Abstract: Responsible AI (RAI) efforts increasingly emphasize the importance of addressing potential harms early in the AI development lifecycle through

safetyarxiv-cs-ai
31 Jul 2026
Safety

AI Security Priorities: A Field-Wide Agenda

DGX agent

arXiv:2607.26069v1 Announce Type: cross Abstract: As AI systems are rapidly integrated into critical economic, governmental, and national security functions, the gap between AI adoption and AI securit

safetyarxiv-cs-ai
31 Jul 2026
Safety

Digital Harf: A Clinically Integrated Multimodal AI System for Pervasive Arabic Speech and Language Therapy

DGX agent

arXiv:2607.27212v1 Announce Type: cross Abstract: Children with Autism Spectrum Disorder in Arabic-speaking countries face compounded barriers to effective speech and language therapy: a shortage of q

safetyarxiv-cs-cl
31 Jul 2026
Agents

(EC)2: Event-Centric Explainability for Cybersecurity Through Multi-Agent LLM Investigations

DGX agent

arXiv:2607.26201v1 Announce Type: cross Abstract: Security operations centers rely on anomaly detection systems to flag suspicious events. Feature-level explanations for anomaly detectors offer limite

agentsarxiv-cs-ai
31 Jul 2026
Model Releases

EMBL AI Librarian: Life-Sciences Knowledge Layer for AI Agents

DGX agent

arXiv:2607.28229v1 Announce Type: new Abstract: The web is increasingly accessed by AI agents rather than humans. Every agent needs knowledge, especially in the life-sciences, where agentic pipelines

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

eta-OPSD: Deriving with Policy Optimization, Training with Self-Distillation

DGX agent

arXiv:2607.28582v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) is a promising approach to improve reasoning language models, but it remains brittle in practice: making it work reli

model-releasesarxiv-cs-lg
31 Jul 2026
Agents

LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger

DGX agent

arXiv:2607.28374v1 Announce Type: new Abstract: Multimodal agents for visual question answering increasingly operate as multi-step trajectories that interleave perception, retrieval, and reasoning, ye

agentsarxiv-cs-lg
31 Jul 2026
Research

Metaphor Tracer: A Theory-Informed Analysis of Hidden States

DGX agent

arXiv:2607.28434v1 Announce Type: cross Abstract: What do a language model's hidden states say about the organization of a single text? From one forward pass, without training, we score every token po

researcharxiv-cs-cl
31 Jul 2026
Research

mmRadarTwin: A Measurement-Calibrated Signal-Level Digital Twin Platform for Indoor mmWave Radar

DGX agent

arXiv:2607.28108v1 Announce Type: new Abstract: Indoor mmWave radar perception is difficult to reproduce because measured range-angle responses depend on scene geometry, material response, multipath,

researcharxiv-cs-cv
31 Jul 2026
Agents

Model-Driven Requirements Configuration with Three-Valued Uncertainty Scoring

DGX agent

arXiv:2607.26220v1 Announce Type: cross Abstract: Context: Large Language Models (LLMs) offer natural-language flexibility for automated requirements elicitation but frequently generate structurally i

agentsarxiv-cs-ai
31 Jul 2026
Model Releases

Models for minimalist RAG: B1ade 335M Embedding and 1B Parameter Small Language Models

DGX agent

arXiv:2607.27506v1 Announce Type: new Abstract: Language and embedding models used in RAG systems are conventionally assumed to require large-scale pretraining and explicit grounding supervision. We p

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Morphological Detection and Classification of Microplastics and Nanoplastics Emerged from Consumer Products by Deep Learning

DGX agent

arXiv:2409.13688v2 Announce Type: replace Abstract: Plastic pollution presents an escalating global issue, impacting health and environmental systems, with micro- and nanoplastics found across mediums

researcharxiv-cs-cv
31 Jul 2026
Model Releases

ORCA-bench: How Ready Are Language Model Agents for Oncall?

DGX agent

arXiv:2607.28545v1 Announce Type: new Abstract: Large language models can write, patch, and search code, but oncall root cause analysis (RCA) demands something different: reasoning over noisy metrics,

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

PoseMaster: A Unified 3D Native Framework for Stylized Pose Generation

DGX agent

arXiv:2506.21076v4 Announce Type: replace Abstract: Pose stylization, which aims to synthesize stylized content aligning with target poses, serves as a fundamental task across 2D, 3D, and video domain

safetyarxiv-cs-cv
31 Jul 2026
Applications

Prompt Chaining in Practice: A Case Study in Automated Scholarly Report Generation

DGX agent

arXiv:2607.27210v1 Announce Type: new Abstract: The exponential growth of scholarly publications requires automated tools for effective information synthesis. However, simple, single-shot prompting me

applicationsarxiv-cs-cl
31 Jul 2026
Agents

RefineSVG: Visual Feedback-Driven Reinforcement Learning for Image-to-SVG Generation

DGX agent

arXiv:2607.27699v1 Announce Type: new Abstract: We propose RefineSVG, a single-step closed-loop visual feedback framework that enables multimodal large language models (MLLMs) to perform high-fidelity

agentsarxiv-cs-cv
31 Jul 2026
Local Ai

ScaFE: Data-Efficient Scar Classification with LLM-Generated Clinical Feature Programs

DGX agent

arXiv:2607.28538v1 Announce Type: new Abstract: Classifying pathological scars from clinical photographs requires distinguishing keloids from hypertrophic scars despite limited expert-labeled data and

local-aiarxiv-cs-cv
31 Jul 2026
Model Releases

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud…

DGX agent

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud agents, world models - general startup / company building f

model-releasesjerry-liu--x
31 Jul 2026
Research

An Attention-Based Framework for Alzheimers Disease Classification Using Resting-State fMRI

DGX agent

arXiv:2607.26746v1 Announce Type: cross Abstract: Accurate identification of Alzheimers disease (AD) using resting-state functional magnetic resonance imaging (rs-fMRI) remains challenging due to the

researcharxiv-cs-lg
30 Jul 2026
Hardware

Can AI agents conduct open-ended AI research? Most evaluations of agents conducting AI research focus on narrow, verifiable tasks. But AI re…

DGX agent

Can AI agents conduct open-ended AI research? Most evaluations of agents conducting AI research focus on narrow, verifiable tasks. But AI research is often open ended. Researchers pick hypotheses, dec

hardwareyann-lecun--x
30 Jul 2026
Agents

Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation

DGX agent

arXiv:2607.26148v1 Announce Type: new Abstract: Autonomous embodied agents must sustain a long decision-making loop that involves perceiving, acting, verifying, and self-correcting over many steps. Cu

agentsarxiv-cs-ro
30 Jul 2026
← Previous
1…6162636465…113
Next →