AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,137 results
Agents

New in Hermes Agent, /compress <topic> to get the compaction model to retain more information on the topic you want it to keep in memory mos…

DGX agent

Nous Research has introduced a new feature in their Hermes Agent system that allows users to use the `/compress ` command to influence how the compaction model prioritizes and retains information. Thi

agentsnous-research--x
12 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

very memgpt / sarah wooders coded. memory isn’t a layer, it is the system. most teams think they’re choosing a model, but they’re really cho…

DGX agent

very memgpt / sarah wooders coded. memory isn’t a layer, it is the system. most teams think they’re choosing a model, but they’re really choosing where their memory lives and like ben thompson says, o

agentsharrison-chase--x
12 Apr 2026
Agents

We’re thrilled to announce @MiniMax_AI M2.7 is now available Day-0 on Fireworks for commercial use. This self-evolving agentic model deliver…

DGX agent

We’re thrilled to announce @MiniMax_AI M2.7 is now available Day-0 on Fireworks for commercial use. This self-evolving agentic model delivers frontier-level performance across: → Software engineering

agentsfireworks-ai--x
12 Apr 2026
Research

Bi-level Heterogeneous Learning for Time Series Foundation Models: A Federated Learning Approach

DGX agent

arXiv:2604.06727v1 Announce Type: new Abstract: Heterogeneity in time series data is more pronounced than in vision or language, as temporal dynamics vary substantially across domains and tasks. Exist

researcharxiv-cs-lg
10 Apr 2026
Research

Development of ML model for triboelectric nanogenerator based sign language detection system

DGX agent

arXiv:2604.06220v1 Announce Type: cross Abstract: Sign language recognition (SLR) is vital for bridging communication gaps between deaf and hearing communities. Vision-based approaches suffer from occ

researcharxiv-cs-ai
10 Apr 2026
Research

DisCEdge: Distributed Context Management for Large Language Models at the Edge

DGX agent

arXiv:2511.22599v2 Announce Type: replace-cross Abstract: Deploying Large Language Model (LLM) services at the edge benefits latency-sensitive and privacy-aware applications. However, the stateless na

researcharxiv-cs-lg
10 Apr 2026
Safety

Event-Level Detection of Surgical Instrument Handovers in Videos with Interpretable Vision Models

DGX agent

arXiv:2604.07577v1 Announce Type: new Abstract: Reliable monitoring of surgical instrument exchanges is essential for maintaining procedural efficiency and patient safety in the operating room. Automa

safetyarxiv-cs-cv
10 Apr 2026
Research

FBS: Modeling Native Parallel Reading inside a Transformer

DGX agent

arXiv:2601.21708v2 Announce Type: replace Abstract: Large language models (LLMs) excel across many tasks, yet inference is still dominated by strictly token-by-token autoregression. Existing accelerat

researcharxiv-cs-ai
10 Apr 2026
Safety

I rest my case: Mythos isn’t AGI. It’s not even better at biology than the last model. It’s tuned to particular things, not a giant advance …

DGX agent

I rest my case: Mythos isn’t AGI. It’s not even better at biology than the last model. It’s tuned to particular things, not a giant advance towards general intelligence. Same as it ever was. @GaryMarc

safetygary-marcus--x
10 Apr 2026
Research

Resource-constrained Amazons chess decision framework integrating large language models and graph attention

DGX agent

arXiv:2603.10512v2 Announce Type: replace Abstract: Artificial intelligence has advanced significantly through the development of intelligent game-playing systems, providing rigorous testbeds for deci

researcharxiv-cs-ai
10 Apr 2026
Local Ai

Toward Personalized Darts Training: A Data-Driven Framework Based on Skeleton-Based Biomechanical Analysis and Motion Modeling

DGX agent

arXiv:2604.01130v3 Announce Type: replace Abstract: As sports training becomes more data-driven, traditional dart coaching based mainly on experience and visual observation is increasingly inadequate

local-aiarxiv-cs-lg
10 Apr 2026
Safety

Towards Privacy-Preserving Large Language Model: Text-free Inference Through Alignment and Adaptation

DGX agent

arXiv:2604.06831v1 Announce Type: cross Abstract: Current LLM-based services typically require users to submit raw text regardless of its sensitivity. While intuitive, such practice introduces substan

safetyarxiv-cs-ai
10 Apr 2026
Tutorials

Transforming the Voice of the Customer: Large Language Models for Identifying Customer Needs

DGX agent

arXiv:2503.01870v2 Announce Type: replace Abstract: Identifying customer needs (CNs) is fundamental to product innovation and marketing strategy. Yet for over thirty years, Voice-of-the-Customer (VOC)

tutorialsarxiv-cs-cl
10 Apr 2026
Safety

VLMShield: Efficient and Robust Defense of Vision-Language Models against Malicious Prompts

DGX agent

arXiv:2604.06502v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face significant safety vulnerabilities from malicious prompt attacks due to weakened alignment during visual integration.

safetyarxiv-cs-lg
10 Apr 2026
Hardware

(1/5) FP4 hardware is here, but 4-bit attention still kills model quality, blocking true end-to-end FP4 serving. To fix that, we propose Att…

DGX agent

(1/5) FP4 hardware is here, but 4-bit attention still kills model quality, blocking true end-to-end FP4 serving. To fix that, we propose Attn-QAT, the first systematic study of quantization-aware trai

hardwarejeremy-howard--x
9 Apr 2026
Tools

A common theme at @aiDotEngineer @swyx @steipete 🦞 Lot’s of fun and meeting great people, my talk about model inference at @superlinked is …

DGX agent

Fardis Makraduli ([@f_makraduli](https://x.com/f_makraduli)) shared a post about attending the AI Engineer ([@aiDotEngineer](https://x.com/aiDotEngineer)) event, noting shared themes and networking...

toolsswyx--x
9 Apr 2026
Tools

At 15:10 today, I’ll be speaking about our SWE-rebench leaderboard at AI Engineer Europe. I'll cover how we build evals and how models cheat…

DGX agent

At 15:10 today, I’ll be speaking about our SWE-rebench leaderboard at AI Engineer Europe. I'll cover how we build evals and how models cheat! Come listen and let's chat! So far, this is the coolest ap

toolsswyx--x
9 Apr 2026
Applications

we got LangPod before GTA6 🙏 this series is gonna be sick - real stuff that breaks with agents, mental models, evals, tooling with some of …

DGX agent

we got LangPod before GTA6 🙏 this series is gonna be sick - real stuff that breaks with agents, mental models, evals, tooling with some of the best builders across industry great to openly share all t

applicationsharrison-chase--x
9 Apr 2026
Model Releases

Abliteration Mitigation via Refusal Aliases

DGX agent

arXiv:2608.18093v1 Announce Type: cross Abstract: Abliteration, the removal of refusal capabilities from large language models by projecting weight matrices orthogonal to an extracted refusal directio

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Cacheable by Design? Training Mixture-of-Experts Routers for Locality Against the Edge Memory-Bandwidth Wall: A Pre-Registered Negative Result with a Systems Measurement Study

DGX agent

arXiv:2608.18261v1 Announce Type: new Abstract: Serving a 235B-parameter Mixture-of-Experts (MoE) model on a single 8 GB GPU is bottlenecked not by compute but by memory bandwidth: decode must stream

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Compress and Forget: bitsandbytes Quantization Amplifies Proactive Interference in LLMs

DGX agent

arXiv:2608.18578v1 Announce Type: new Abstract: Proactive interference (PI) is a documented failure mode in large language models in which retrieval of a repeatedly overwritten value degrades as prior

model-releasesarxiv-cs-cl
20 Aug 2026
Model Releases

FlashAttention for Scalable Vector Architectures

DGX agent

arXiv:2608.18656v1 Announce Type: new Abstract: Inference with transformer models on CPUs is increasingly important, especially for Small Language Models (SLMs), where vector architectures are emergin

model-releasesarxiv-cs-lg
20 Aug 2026
Model Releases

LACONIC: Dense-Level Effectiveness for Scalable Sparse Retrieval via a Two-Phase Training Curriculum

DGX agent

arXiv:2601.01684v2 Announce Type: replace-cross Abstract: While dense retrieval models have been the standard for state-of-the-art information retrieval, their deployment is often constrained by high

model-releasesarxiv-cs-cl
20 Aug 2026
Model Releases

Measuring the Partial-Credit Gap: A Strict Benchmark on Vietnam's 2025 Convex Marking Scheme

DGX agent

arXiv:2608.18336v1 Announce Type: new Abstract: When evaluating language models on human exams, benchmarks typically score each response as right or wrong and report the overall accuracy. This approac

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

OmniHandwritingOCR: A Diagnostic Benchmark for Evaluating Multimodal LLMs in Handwritten OCR Scenarios

DGX agent

arXiv:2608.18586v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used as OCR systems in document and knowledge-processing pipelines, but their ability to fai

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Qwen 3.8 27B KV f16 vs q8_0 are not equivalents

DGX agent

I'm testing it since release, now with UD 3.0 in my AMD R9700 with ROCm, I always read everywhere that F16 and q8_0 for KV cache are essentially the same... well, I tested it and I can see differences

model-releasesr-localllama
20 Aug 2026
Research

Reconstruction of continuum robots by marker-free shape registration of image data using a kinematic model

DGX agent

arXiv:2405.15336v2 Announce Type: replace Abstract: Continuum robots are slender, flexible manipulators that navigate confined, curved workspaces and are gaining traction in aerospace, inspection, aut

researcharxiv-cs-ro
20 Aug 2026
Hardware

Sanja Fidler’s world model startup Veeda AI raises $90M in seed funding

DGX agent

Veeda AI, a startup led by a team of former Nvidia Corp. researcher and renowned computer scientist Sanja Fidler, has taken its bow on the main stage after raising 90 million in a seed funding round t

hardwaresiliconangle
20 Aug 2026
Research

Temporal Multi-Signal Fusion for Token-Level Hallucination Detection

DGX agent

arXiv:2608.18115v1 Announce Type: cross Abstract: Token-level hallucination detectors score each token independently from a single signal, and fail exactly when the generating model is confidently wro

researcharxiv-cs-ai
20 Aug 2026
Model Releases

TinySearch v0.6.1 - still a lightweight web research tool for local LLMs, now with bring-your-own-browser support

DGX agent

Hey everyone, Posted TinySearch here a few versions ago and got a bunch of useful feedback, so figured I'd post an update because the thing has changed quite a bit since then. Repo: [https://github.co

model-releasesr-localllama
20 Aug 2026
Hardware

Tuning the Stochastic Machine: A Systems Engineer's Operating Model for Human-AI Engineering

DGX agent

arXiv:2608.19125v1 Announce Type: new Abstract: When an expert corrects an LLM assistant's error, the correction usually dies with the session, and the error class returns. I argue this is an operatio

hardwarearxiv-cs-ai
20 Aug 2026
Model Releases

A Residual Learning Approach for Unsteady Aerodynamic Load Prediction

DGX agent

arXiv:2608.17894v1 Announce Type: cross Abstract: This paper investigates the feasibility of using residual learning to improve unsteady aerodynamic load prediction for aeroelastic applications. The m

model-releasesarxiv-cs-lg
19 Aug 2026
Research

An Investigation of Translationese in the Generations of Multilingual Large Language Models

DGX agent

arXiv:2608.17399v1 Announce Type: new Abstract: Text which has been translated from another language tends to carry with it evidence of translationnicode{x2014}hence, it is often referred to as extit{

researcharxiv-cs-cl
19 Aug 2026
Local Ai

AntLing’ve open-sourced 6 Base Model checkpoints for Ling-3.0-tiny & Ling-3.0-flash, covering pre-trained, mid-trained, and WSM-merged stages.

DGX agent

None has undergone post-training, giving researchers flexible starting points for continued pre-training, fine-tuning, and further research. Two key highlights: - They use WSM to replace LR decay with

local-air-localllama
19 Aug 2026
Research

ChatPlanner: A Large Language Model Framework for Personalized Public Transit Routing

DGX agent

arXiv:2606.15315v2 Announce Type: replace Abstract: Personalized public transit routing in public transit systems remains challenging due to the difficulty of capturing and integrating diverse user pr

researcharxiv-cs-ai
19 Aug 2026
Research

Cluster Aggregated GAN (CAG): A Cluster-Based Hybrid Model for Appliance Pattern Generation

DGX agent

arXiv:2512.22287v4 Announce Type: replace-cross Abstract: Synthetic appliance data are essential for developing non-intrusive load monitoring algorithms and enabling privacy preserving energy research

researcharxiv-cs-ai
19 Aug 2026
Model Releases

Do LLMs Know a Good Hypothesis When They See One? Logit-Based Energy Scoring Outperforms Prompted LLM-as-Judge for Scientific Hypothesis Ranking

DGX agent

arXiv:2608.17270v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for scientific hypothesis generation. However, evaluating generated hypotheses remains a challenge fo

model-releasesarxiv-cs-ai
19 Aug 2026
Model Releases

Neural Operator-Based Nonlinear Nudging for Chaotic Dynamical Systems

DGX agent

arXiv:2508.05778v2 Announce Type: replace Abstract: Nudging is an empirical data assimilation technique that incorporates an observation-driven control term into the model dynamics. The trajectory of

model-releasesarxiv-cs-lg
19 Aug 2026
Agents

PDDL-ART: Autonomous Symbolic Abstraction From Demonstration For Long-Horizon Robotic Manipulation Using Vision-Language Models

DGX agent

arXiv:2608.17146v1 Announce Type: new Abstract: Symbolic planning with PDDL offers a principled framework for long-horizon robot manipulation, but constructing accurate PDDL domain and problem descrip

agentsarxiv-cs-ro
19 Aug 2026
Model Releases

Same GRPO recipe on three from-scratch LLMs (353M/316M/672M) gave three different outcomes, with no clean relationship to scale [P]

DGX agent

I trained three LLMs from scratch in raw PyTorch then post-trained each one with SFT and then GRPO. Same process every time: same synthetic arithmetic curriculum, same reward function, same hyperparam

model-releasesr-machinelearning
19 Aug 2026
Model Releases

TileMix: Tile-Centric Mixed-Precision Attention for LLM Inference Acceleration

DGX agent

arXiv:2608.17336v1 Announce Type: new Abstract: Long-context prefill in large language models (LLMs) incurs substantial computation and memory traffic because dense self-attention computes quadratic q

model-releasesarxiv-cs-ai
19 Aug 2026
Model Releases

What Aggregate Scores Miss: Measuring Item-Level Regressions in Commercial LLM API Migrations

DGX agent

arXiv:2608.17719v1 Announce Type: cross Abstract: Context: Software systems that depend on commercial large language model APIs must migrate to successor versions when vendors deprecate older models.

model-releasesarxiv-cs-ai
19 Aug 2026
Safety

Audio-Visual Segmentation via Depth-Guided Collaborative Modeling

DGX agent

arXiv:2608.16285v1 Announce Type: cross Abstract: Audio-Visual Segmentation (AVS) is a fundamental task in multimodal perception that performs pixel-level segmentation of sounding objects in videos by

safetyarxiv-cs-ai
18 Aug 2026
Applications

Beyond Binary Priorities: Multi-Tier SLA Scheduling for Large Language Model Serving

DGX agent

arXiv:2608.16336v1 Announce Type: cross Abstract: Modern LLM serving deployments must simultaneously satisfy heterogeneous service-level objectives (SLOs) across a diverse population of user tiers, ra

applicationsarxiv-cs-lg
18 Aug 2026
Hardware

BrainLinear: A Linear Model for Brain Network Analysis in Sparse Tangent Subspaces

DGX agent

arXiv:2608.15266v1 Announce Type: cross Abstract: Functional connectome analysis examines brain-region interactions to understand and identify disorders such as autism spectrum disorder and Alzheimer'

hardwarearxiv-cs-lg
18 Aug 2026
Model Releases

Building operational resilience with agentic AI in financial services

DGX agent

For financial institutions, operational resilience has long been embedded in regulatory and supervisory expectations — to say nothing of the high expectations of consumers. With the implementation of

model-releasesgoogle-cloud-ai
18 Aug 2026
Research

Coarse-to-Fine Multi-Resolution Diffusion Models for Trajectory Generation in Urban Systems

DGX agent

arXiv:2608.14570v1 Announce Type: new Abstract: Understanding human mobility is critical for a wide range of urban applications, including traffic management, epidemic control, and urban planning. How

researcharxiv-cs-lg
18 Aug 2026
Model Releases

Evaluating Agentic Code Repair Capabilities in Distributed Systems

DGX agent

arXiv:2608.14863v1 Announce Type: cross Abstract: LLM-based coding agents have advanced rapidly on single-process SWE tasks, with frontier models now clustering in the high-70s on SWE-bench Verified.

model-releasesarxiv-cs-ai
18 Aug 2026
← Previous
1…291292293294295…1316
Next →