AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “concepts”

GridTimelineEvolution
2,542 results
11 Aug 2026

Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family

SafetyDGX agent

arXiv:2604.05971v2 Announce Type: replace-cross Abstract: Recent research has shown that contrastive vision-language models such as CLIP often lack fine-grained understanding of visual content. While

LGNNIC: Acceleration of Large-Scale GNN Training using SmartNICs

Model ReleasesDGX agent

arXiv:2608.07733v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are widely used across domains such as natural sciences, social network analysis, chip design, and recommendation systems

Lying mirror using structured surfaces

ResearchDGX agent

arXiv:2410.15521v2 Announce Type: replace-cross Abstract: We introduce an all-optical system, termed the 'lying mirror', to hide input information by transforming it into misleading, ordinary-looking

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Machine Learning and Data Analysis Using Posets: A Survey

SafetyDGX agent

arXiv:2404.03082v3 Announce Type: replace Abstract: Partially ordered sets (posets) are discrete mathematical structures that formalize the notion of comparison without forcing every pair of objects t

MiniMax-H3: ~38 GB less VRAM with Runtime LoRA Bypass — DoRA Dynamic LoRA Loader v1.0.39

Local AiDGX agent

GitHub: https://github.com/xmarre/ComfyUI-DoRA-Dynamic-LoRA-Loader Release v1.0.39: https://github.com/xmarre/ComfyUI-DoRA-Dynamic-LoRA-Loader/releases/tag/v1.0.39 Also available through ComfyUI Manag

Private Etymology: Designing Relational Reuse of Shared Symbols in Long-Term Human-AI Interaction

Local AiDGX agent

arXiv:2608.08443v1 Announce Type: cross Abstract: Previous studies have shown that people can develop shared symbols, partner-specific expressions, personal idioms, inside jokes, and other parts of a

Recurrent Neural Networks Beyond Time: Learning from Multiple Ordered Projections

ResearchDGX agent

arXiv:2608.09690v1 Announce Type: new Abstract: Recurrent neural networks (RNNs) are widely used for sequence learning, yet their application is commonly associated with temporal data, although recurr

SAKE: Structured Agentic Knowledge Extrapolation for Complex LLM Reasoning via Reinforcement Learning

AgentsDGX agent

arXiv:2505.15062v5 Announce Type: replace-cross Abstract: Knowledge extrapolation is the process of inferring novel information by combining and extending existing knowledge that is explicitly availab

The Spectral Neuron

Local AiDGX agent

arXiv:2608.08003v1 Announce Type: cross Abstract: As machine learned models increase in complexity and expressive power, features of simpler models, such as interpretability and control over the shape

VCU-Bridge: Hierarchical Visual Connotation Understanding via Semantic Bridging

Model ReleasesDGX agent

arXiv:2511.18121v2 Announce Type: replace-cross Abstract: While Multimodal Large Language Models (MLLMs) excel on benchmarks, their processing paradigm differs from the human ability to integrate visu

10 Aug 2026

A foundation-model approach to pediatric headache classification from rs-fMRI

ResearchDGX agent

arXiv:2608.07287v1 Announce Type: new Abstract: Headache is the most common neurological disorder in children and substantially affects quality of life. We investigated whether resting-state functiona

Agentic AI: User Empowerment or Enclosure?

AgentsDGX agent

arXiv:2608.06510v1 Announce Type: cross Abstract: Agentic AI promises a more flexible form of digital agency: systems that can act on users' behalf, from filtering content to negotiating prices to sel

Cascading Through the Hierarchy: Regularizer-Induced Feature Detection as Phase Transitions in Deep Linear Neural Networks

Model ReleasesDGX agent

arXiv:2608.06597v1 Announce Type: cross Abstract: A scientific theory of deep learning, comprising learning dynamics and statistical properties of learned models, is rapidly gaining attention. One of

DREAMS: Diverse Reactions of Engagement and Attention Mind States Dataset

ResearchDGX agent

arXiv:2608.06382v1 Announce Type: cross Abstract: Active attention and engagement are important in improving users' learning experiences. Engagement refers to the level of involvement and interest ind

Dual-Node NVIDIA DGX Spark over Tailscale: A Remote-Access Testbed for Distributed LLM Training and Cyber-Threat-Intelligence Fine-Tuning

Local AiDGX agent

arXiv:2608.07226v1 Announce Type: cross Abstract: Compact AI systems make local language-model experimentation increasingly accessible, yet practical evidence for multi-node training on desktop-class

Equivariant Sparse Autoencoders: Mechanistic Interpretability of Neural Networks on Symmetric Data

ApplicationsDGX agent

arXiv:2511.09432v2 Announce Type: replace Abstract: Machine learning (ML) models achieve remarkable performance but remain hard to interpret due to their scale and complexity. In particular, their act

How Malachyte solves retail’s cold-start problem with managed real-time AI

Local AiDGX agent

What’s the best way to recommend products to little-known users? We’ve spent our careers trying to solve this problem for major companies like Spotify and Priceline, and it’s why Sidd founded Malachyt

Human-AI Perceptual Alignment by Playing Hues and Cues

Local AiDGX agent

arXiv:2608.07141v1 Announce Type: new Abstract: Evaluating the perceptual alignment between Contrastive Vision-Language Models (CVLMs) and humans is typically constrained by traditional benchmarks tha

Human-Centered Explainable AI for TinyML Edge Devices: A Pareto-Based Selection Framework with LLM-Guided Design

Local AiDGX agent

arXiv:2608.07091v1 Announce Type: cross Abstract: Edge Artificial Intelligence (Edge AI) enables the deployment of AI models directly on local edge devices, while such deployments are subject to stric

Latent Fact-Checking: Detecting Misinformation through Activation Engineering

Model ReleasesDGX agent

arXiv:2608.06417v1 Announce Type: cross Abstract: The proliferation of misinformation online has driven demand for scalable detection systems. While most existing approaches rely on surface-level ling

MI-MIDI: Mechanistic Interpretability of Text-to-MIDI Generation Models via Probing, Lenses and Steering

Model ReleasesDGX agent

arXiv:2608.06638v1 Announce Type: cross Abstract: Mechanistic interpretability of music generation has concentrated on audio models, leaving symbolic models largely unexplored. We analyze two public t

MiGHT-EHR: A Multi-task Graph Transformer for Heterogeneous Temporal Electronic Health Records

ResearchDGX agent

arXiv:2608.06430v1 Announce Type: new Abstract: Learning from Electronic Health Records (EHRs) has gained significant attention due to its potential to improve clinical prediction. However, effective

MiniMax H3 with a 4B or 8B text encoder instead of the 32B: update, the voice matches now

Local AiDGX agent

MiniMax H3 loads a 32B text encoder, 15.7 GB, just to turn your prompt into a conditioning tensor. I replaced it with a Qwen3-VL 4B or 8B plus a learned map into the same space. Same DiT, same VAEs, s

Model Confidence Under Answer-Preserving Attacks: An Informativeness-Manipulability Frontier

Model ReleasesDGX agent

arXiv:2608.06571v1 Announce Type: cross Abstract: Deployed vision-language systems often gate their answers on confidence, making confidence robustness relevant to oversight. We study confidence reado

Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery

Model ReleasesDGX agent

arXiv:2608.06931v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in scientific discovery, yet it remains unclear whether they can support complex real laboratory

Seeking SOTA: Time-Series Forecasting Must Adopt Taxonomy-Specific Evaluation to Dispel Illusory Gains

Model ReleasesDGX agent

arXiv:2603.15506v2 Announce Type: replace-cross Abstract: We argue that the current practice of evaluating AI/ML time-series forecasting models, predominantly on benchmarks characterized by strong, pe

Social World Models

Model ReleasesDGX agent

arXiv:2509.00559v3 Announce Type: replace Abstract: Humans intuitively navigate social interactions by simulating unspoken dynamics and reasoning about others' perspectives, even with limited informat

9 Aug 2026

GitHub Models is now retired

Model ReleasesDGX agent

GitHub Models is now retired I missed this news until today, when the GitHub Actions run for my simonw/research repository failed with this error message: GitHub Models is temporarily unavailable as p

8 Aug 2026

Showoff Saturday: Local 4x 6000 Pro (multi-year progression)

Model ReleasesDGX agent

Not the biggest or shiniest, but it's mine From gaming machine inference on the original llama models, to a 4x RTX 6000 Pro Max Q + 4x 3090s local AI cluster. Pictures are in reverse chronological ord

7 Aug 2026

Analogy as Nonparametric Bayesian Inference over Relational Systems

TutorialsDGX agent

arXiv:2006.04156v2 Announce Type: replace Abstract: Our inferences in the real world are rarely naive - we acquire experiences through our lifetime that can help us more quickly understand the structu

BlockPython: A Process-Aware Agent-Supported Platform for the Transition from Block-Based to Python Programming

AgentsDGX agent

arXiv:2608.05716v1 Announce Type: new Abstract: The transition from block-based to text-based programming requires learners to convert visible program structures into abstract textual expressions, whi

DARAD: Dual Adapters and Ranking-Aware Distillation for Continual Remote Sensing Image-Text Retrieval

SafetyDGX agent

arXiv:2608.06059v1 Announce Type: new Abstract: With the rapid growth of Earth observation technologies, remote sensing archives are rapidly expanding, making remote sensing image-text retrieval (RS-I

Evaluating and Improving Pedagogical Fit in LLM-Based AI Tutors with the Pedagogical Suitability Index

Model ReleasesDGX agent

arXiv:2608.05411v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as AI tutors, but a correct answer is not always a pedagogically appropriate one. In classroom learni

Evidential Rule Learning for Interpretable Classification with Abstention

Model ReleasesDGX agent

arXiv:2608.05859v1 Announce Type: cross Abstract: Interpretable classification often requires more than accurate predictions for real-life deployment: models should be transparent about the evidence b

Got job as Director of AI and Systems development self-taught

Model ReleasesDGX agent

Hey everyone, I just wanted to share my journey here for some motivation. Three years ago, I saw the sudden spike in AI and realized it was the future of tech. My goal at the time was to be an indie g

Sliding Sensors: Configurable Confidence in State Estimation for Continuum Robots

ResearchDGX agent

arXiv:2608.05410v1 Announce Type: new Abstract: Continuum robots often operate in uncertain environments, where accurate state estimation is essential for safe interactions. Estimate uncertainty is in

You can play it here: https://simonw.github.io/raccoon-heist-codex/ For comparison, here's Fable 5 + Claude Code's game, built from the exac…

Model ReleasesDGX agent

You can play it here: https://simonw.github.io/raccoon-heist-codex/ For comparison, here's Fable 5 + Claude Code's game, built from the exact same prompt https://x.com/simonw/status/208508951822360205

6 Aug 2026

Agentic Future Ready With BigQuery: Continually Improving Price-Performance, Zero Effort

Model ReleasesDGX agent

In the modern data landscape, query performance tuning and managing system price-performance is challenging, especially as the number of agentic workloads increase. Even for experienced developers and

An entropic explanation of insistence on sameness in autism

TutorialsDGX agent

arXiv:2608.04616v1 Announce Type: new Abstract: An information theory-based framework is proposed in attempt to explain insistence on sameness in autism as an instance of a general behavior pattern in

Approximate Multi-Objective Search Under Rulebooks

SafetyDGX agent

arXiv:2608.04398v1 Announce Type: cross Abstract: Robotic planning often involves multiple objectives with complex priority relationships, such as safety, efficiency, and regulatory compliance. Rulebo

Do LLMs Know What Is Private Internally? Probing and Steering Contextual Privacy Norms in Large Language Model Representations

ResearchDGX agent

arXiv:2604.00209v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in high-stakes settings, yet they frequently violate contextual privacy by disclosing private

From Score Matrices to Football-Aware Match-State Simulation: An Auditable LLM Harness for Exact-Score Reranking

Model ReleasesDGX agent

arXiv:2608.05030v1 Announce Type: new Abstract: Football score forecasting combines a strong statistical core with a difficult contextual edge. Dynamic Poisson-family models estimate team strength, ex

Improving Auto-Design of Neural PDE Solvers with a Domain-Specific Language

AgentsDGX agent

arXiv:2608.04384v1 Announce Type: new Abstract: Neural PDE solver auto-design is fundamentally a search-space representation problem. In the space of unrestricted Python programs, valid solvers form a

IslamicTurathBench: A Multi-Task, Multi-Discipline Benchmark for Evaluating Large Language Models on the Islamic Scholarly Tradition (turath)

Model ReleasesDGX agent

arXiv:2608.04703v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for question answering, education, and research, including in religious and cultural domains where an

NeuMoSync: End-to-End Neuromodulatory Control for Plasticity and Adaptability in Continual Learning

TutorialsDGX agent

arXiv:2608.04358v1 Announce Type: new Abstract: Continual learning (CL) requires models to learn tasks sequentially, yet deep neural networks often suffer from plasticity loss and poor knowledge trans

Regularization can make diffusion models more efficient

ResearchDGX agent

arXiv:2502.09151v3 Announce Type: replace Abstract: Diffusion models are one of the key architectures of generative AI. Their main drawback, however, is the computational costs. This study indicates t

5 Aug 2026

A game theory for foundation models shows new paths to rational cooperation through similarity inference

SafetyDGX agent

arXiv:2608.03958v1 Announce Type: new Abstract: As autonomous agents powered by foundation models are increasingly integrated into social and economic systems, understanding the principles governing t

Activation-Guided Neuron Intervention to Induce Alzheimer's-Related Computational Language Phenotypes in a Large Language Model

ResearchDGX agent

arXiv:2608.03067v1 Announce Type: new Abstract: Changes in spontaneous speech provide an early signal of cognitive dysfunction in Alzheimer's disease (AD) that large language models (LLMs) can detect.

Adversarial Fast-Moving Real-World Domains as Test Beds for Benchmarking AI Scientist Capabilities

Model ReleasesDGX agent

arXiv:2608.03569v1 Announce Type: new Abstract: Benchmarking the ability of AI scientists to generate novel ideas is notoriously difficult. Existing benchmarks in this field have made progress in eval

ArtECulture: Benchmarking Culture-Conditioned Visual Emotion Understanding in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2608.03358v1 Announce Type: new Abstract: Existing visual emotion understanding methods typically ignore cultural variations in emotional perception. We introduce culture-conditioned visual emot

Assessing the Effect of Cross-Domain Mapping on Creativity in Humans and Large Language Models

ResearchDGX agent

arXiv:2603.19087v2 Announce Type: replace Abstract: Creativity is the ability to come up with novel ideas, a capacity crucial for human development and flourishing. Are large language models (LLMs) cr

Autoreflection: How Agentic Strange Loops Turn Human Culture into AI Infrastructure

AgentsDGX agent

arXiv:2608.03800v1 Announce Type: cross Abstract: An LLM-based agent is a loop that reads itself. Agentic frameworks externalize identity, memory, and disposition into editable files. The agent loads

Beyond Either-Or Reasoning: Transduction and Induction as Cooperative Problem-Solving Paradigms

Model ReleasesDGX agent

arXiv:2505.14744v3 Announce Type: replace-cross Abstract: Traditionally, in Programming-by-example (PBE) the goal is to synthesize a program from a small set of input-output examples. Lately, PBE has

Chat Debugging: An Exploratory Study of Human-AI Collaboration to Debug Analog Circuits

ResearchDGX agent

arXiv:2608.02955v1 Announce Type: cross Abstract: This research paper describes an exploratory study on the effectiveness of Chat Debugging: troubleshooting malfunctioning analog circuits on breadboar

Dr. AGENTONOMICS: A Didactic Experiment of AGENTONOMICS

AgentsDGX agent

arXiv:2608.03524v1 Announce Type: new Abstract: AGENTONOMICS is a framework that treats AI agents as economic entities that can be designed, managed, and governed through an integrated management arch

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners

Model ReleasesDGX agent

arXiv:2608.03206v1 Announce Type: cross Abstract: Large language models (LLMs) power educational applications from tutoring to essay scoring, but each is a point solution to a single task, and only re

... four years later, I got the new Claude Fable 5 to actually build the game https://x.com/simonw/status/2085089518223602058

Model ReleasesDGX agent

... four years later, I got the new Claude Fable 5 to actually build the game https://x.com/simonw/status/2085089518223602058 Four years ago today I tweeted about having GPT-3 and DALL-E come up with

ICO: Enhancing Semantic-Shift Jailbreaks via Iterative Context Optimization

SafetyDGX agent

arXiv:2608.03210v1 Announce Type: new Abstract: Foundation models have achieved remarkable success across diverse tasks, but they remain vulnerable. To investigate such vulnerabilities, semantic-shift

Learning Biomechanically Plausible Human Motion from Sparse Radar Point Clouds

ResearchDGX agent

arXiv:2608.03637v1 Announce Type: new Abstract: Radar-based human pose estimation has focused on improving learning algorithms while representing the body as unconstrained keypoint coordinates. We add

LLMs Can Annotate Attribution Graphs

ResearchDGX agent

arXiv:2608.02632v1 Announce Type: new Abstract: Circuit tracing is an exciting technique for revealing the internal computation of language models, but it requires a time-intensive manual step of grou

← Previous
1…2021222324…43
Next →