AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Model Releases

Making Room for AI: Multi-GPU Molecular Dynamics with Deep Potentials in GROMACS

DGX agent

arXiv:2604.07276v1 Announce Type: cross Abstract: GROMACS is a de-facto standard for classical Molecular Dynamics (MD). The rise of AI-driven interatomic potentials that pursue near-quantum accuracy a

model-releasesarxiv-cs-ai
10 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

MedRoute: RL-Based Dynamic Specialist Routing in Multi-Agent Medical Diagnosis

DGX agent

arXiv:2604.06180v1 Announce Type: cross Abstract: Medical diagnosis using Large Multimodal Models (LMMs) has gained increasing attention due to capability of these models in providing precise diagnose

agentsarxiv-cs-lg
10 Apr 2026
Safety

MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning

DGX agent

arXiv:2604.08203v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) hold immense promise for complex clinical tasks, but their reasoning capabilities are often constrained by text-on

safetyarxiv-cs-cv
10 Apr 2026
Research

MICA: Multivariate Infini Compressive Attention for Time Series Forecasting

DGX agent

arXiv:2604.06473v1 Announce Type: new Abstract: Multivariate forecasting with Transformers faces a core scalability challenge: modeling cross-channel dependencies via attention compounds attention's q

researcharxiv-cs-lg
10 Apr 2026
Applications

Mitigating Domain Drift in Multi Species Segmentation with DINOv2: A Cross-Domain Evaluation in Herbicide Research Trials

DGX agent

arXiv:2508.07514v3 Announce Type: replace Abstract: Reliable plant species and damage segmentation for herbicide field research trials requires models that can withstand substantial real-world variati

applicationsarxiv-cs-cv
10 Apr 2026
Agents

NaviSplit: Dynamic Multi-Branch Split DNNs for Efficient Distributed Autonomous Navigation

DGX agent

arXiv:2406.13086v2 Announce Type: replace Abstract: Lightweight autonomous unmanned aerial vehicles (UAV) are emerging as a central component of a broad range of applications. However, autonomous navi

agentsarxiv-cs-ro
10 Apr 2026
Safety

Neural Computers

DGX agent

arXiv:2604.06425v1 Announce Type: cross Abstract: We propose a new frontier: Neural Computers (NCs) -- an emerging machine form that unifies computation, memory, and I/O in a learned runtime state. Un

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Neural parametric representations for thin-shell shape optimisation

DGX agent

arXiv:2604.06612v1 Announce Type: cross Abstract: Shape optimisation of thin-shell structures requires a flexible, differentiable geometric representation suitable for gradient-based optimisation. We

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Physics-Informed Neural Networks for Joint Source and Parameter Estimation in Advection-Diffusion Equations

DGX agent

arXiv:2512.07755v3 Announce Type: replace-cross Abstract: Recent studies have demonstrated the success of deep learning in solving forward and inverse problems in engineering and scientific computing

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Possible memory leak in Ollama when using Claude Code?

DGX agent

Users in the r/ollama community have reported a possible memory leak occurring in Ollama when it is used as a backend with Claude Code, with Ollama runner processes not always being properly termin...

model-releasesr-ollama
10 Apr 2026
Safety

Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning

DGX agent

arXiv:2505.24499v2 Announce Type: replace Abstract: Generating high-quality Scalable Vector Graphics (SVGs) is challenging for Large Language Models (LLMs), as it requires advanced reasoning for struc

safetyarxiv-cs-cv
10 Apr 2026
Research

SAT: Balancing Reasoning Accuracy and Efficiency with Stepwise Adaptive Thinking

DGX agent

arXiv:2604.07922v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have revolutionized complex problem-solving, yet they exhibit a pervasive 'overthinking', generating unnecessarily long

researcharxiv-cs-cl
10 Apr 2026
Model Releases

SceneScribe-1M: A Large-Scale Video Dataset with Comprehensive Geometric and Semantic Annotations

DGX agent

arXiv:2604.07990v1 Announce Type: new Abstract: The convergence of 3D geometric perception and video synthesis has created an unprecedented demand for large-scale video data that is rich in both seman

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Selective Neuron Amplification for Training-Free Task Enhancement

DGX agent

arXiv:2604.07098v1 Announce Type: new Abstract: Large language models often fail on tasks they seem to already understand. In our experiments, this appears to be less about missing knowledge and more

researcharxiv-cs-lg
10 Apr 2026
Applications

SELFDOUBT: Uncertainty Quantification for Reasoning LLMs via the Hedge-to-Verify Ratio

DGX agent

arXiv:2604.06389v1 Announce Type: new Abstract: Uncertainty estimation for reasoning language models remains difficult to deploy in practice: sampling-based methods are computationally expensive, whil

applicationsarxiv-cs-ai
10 Apr 2026
Model Releases

SentinelSphere: Integrating AI-Powered Real-Time Threat Detection with Cybersecurity Awareness Training

DGX agent

arXiv:2604.06900v1 Announce Type: cross Abstract: The field of cybersecurity is confronted with two interrelated challenges: a worldwide deficit of qualified practitioners and ongoing human-factor wea

model-releasesarxiv-cs-ai
10 Apr 2026
Local Ai

Should we be optimizing for limited compute instead of more parameters? Thoughts?

DGX agent

"The search did not return the specific Reddit thread. However, I can provide a summary based on what the topic is broadly about within the local-AI/Ollama community context:

local-air-ollama
10 Apr 2026
Model Releases

TGIF! Here are some of our favorite updates from the past week: — Notebooks in @GeminiApp, an integration with @NotebookLM that enables you …

DGX agent

TGIF! Here are some of our favorite updates from the past week: — Notebooks in @GeminiApp, an integration with @NotebookLM that enables you to retrieve context from your private notebooks or convert y

model-releasesgoogle-ai--x
10 Apr 2026
Safety

The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training

DGX agent

arXiv:2604.07754v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) raises significant ethical and safety concerns. While LLM alignment techniques are adopted to improve m

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

The Geometry of Forgetting

DGX agent

arXiv:2604.06222v1 Announce Type: cross Abstract: Why do we forget? Why do we remember things that never happened? The conventional answer points to biological hardware. We propose a different one: ge

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Verify Before You Commit: Towards Faithful Reasoning in LLM Agents via Self-Auditing

DGX agent

arXiv:2604.08401v1 Announce Type: cross Abstract: In large language model (LLM) agents, reasoning trajectories are treated as reliable internal beliefs for guiding actions and updating memory. However

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Grok Law

DGX agent

Grok Law Grok-4.20 just ranked #1 in Legal & Government on Chatbot Arena It’s officially outperforming Anthropic’s Opus 4.6 and Google’s Gemini 3.1 Pro Grok is actively helping people navigate real la

model-releaseselon-musk--x
9 Apr 2026
Agents

'Harnesses are intimately tied to memory, which means that by choosing an open harness you are choosing to own your memory, and not have it …

DGX agent

'Harnesses are intimately tied to memory, which means that by choosing an open harness you are choosing to own your memory, and not have it be locked into a proprietary harness or tied to a single mod

agentsharrison-chase--x
9 Apr 2026
Model Releases

Introducing Gemma 4 31B from @GoogleDeepMind on Together AI. AI natives can now use Gemma 4 31B on Together and benefit from reliable infere…

DGX agent

Introducing Gemma 4 31B from @GoogleDeepMind on Together AI. AI natives can now use Gemma 4 31B on Together and benefit from reliable inference for multimodal reasoning, tool use, and agentic workflow

model-releasestogether-ai--x
9 Apr 2026
Model Releases

Sparks unicorn https://x.com/emollick/status/2024756029121020236?s=20

DGX agent

Sparks unicorn https://x.com/emollick/status/2024756029121020236?s=20 Here is the Gemini 3.1 'Sparks unicorn' (This is created using TikZ, which is a language built for scientific diagrams & very much

model-releasesethan-mollick--x
9 Apr 2026
Research

We should view the history of physics as a long-running program synthesis task. Kepler and Newton were searching the space of possible symbo…

DGX agent

We should view the history of physics as a long-running program synthesis task. Kepler and Newton were searching the space of possible symbolic models to find the simplest one that would best satisfy

researchfrancois-chollet--x
9 Apr 2026
Safety

Again, if you care about computer security, read the red team report: https://red.anthropic.com/2026/mythos-preview/

DGX agent

Anthropic's Frontier Red Team report (April 2026) details the cybersecurity capabilities of Claude Mythos Preview, a general-purpose frontier model that performs strongly across the board but is s...

safetyethan-mollick--x
8 Apr 2026
Industry

pay for opus to write slop code pay for mythos to fix slop code

DGX agent

Emad Mostaque (founder of Stability AI) posted a sardonic observation on X highlighting an ironic dynamic in the AI coding market: users pay for Claude Opus to generate low-quality 'slop' code, the...

industryemad-mostaque--x
8 Apr 2026
Safety

this is interesting. 1. Did Anthropic forget to run a control? 2. Where does this leave us?

DGX agent

this is interesting. 1. Did Anthropic forget to run a control? 2. Where does this leave us? New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped anal

safetygary-marcus--x
8 Apr 2026
Local Ai

Today, we are launching our collaboration with @nomic_ai to make AI agents more effectively and efficiently understand complex PDF documents…

DGX agent

Today, we are launching our collaboration with @nomic_ai to make AI agents more effectively and efficiently understand complex PDF documents. Nomic's new nomic-layout-v1 model allows your AI agents to

local-ainomic-ai--x
8 Apr 2026
Model Releases

Trying to DIY your own document parser by screenshotting into a frontier VLM (Opus, 5.4, Gemini) carries when you try to scale it up into pr…

DGX agent

Trying to DIY your own document parser by screenshotting into a frontier VLM (Opus, 5.4, Gemini) carries when you try to scale it up into production workflows. Here are two edge cases we've observed:

model-releasesjerry-liu--x
8 Apr 2026
Agents

Using a non-hermetic agent and having trouble with GPT 5.4 or your Codex sub? The rumors are true: it works a lot better with the Hermes spe…

DGX agent

Nous Research's **Hermes Agent** is an open-source agentic framework by NousResearch that delivers significantly improved reliability when using GPT-5.x or Codex (OpenAI Codex subscription) models ...

agentsnous-research--x
8 Apr 2026
Model Releases

Writing fiction seems to be a genuine weak spot for LLMs that is not improving as rapidly as almost every other area. There may be a lot of …

DGX agent

Writing fiction seems to be a genuine weak spot for LLMs that is not improving as rapidly as almost every other area. There may be a lot of reasons why this is happening. It would be a really interest

model-releasesethan-mollick--x
8 Apr 2026
Agents

Come check out GLM 5.1 in Code Arena for agentic web development tasks using tools. Don’t forget to vote, Code Arena scores are coming up ne…

DGX agent

GLM-5.1 is Zhipu AI's next-generation flagship model for agentic engineering, achieving state-of-the-art performance on SWE-Bench Pro and leading its predecessor GLM-5 by a wide margin on NL2Repo (...

agentszhipu-ai--x
7 Apr 2026
Model Releases

“gpt2-large is too powerful to be publicly released” vibes

DGX agent

Julien Chaumond (co-founder of Hugging Face) posted a tweet referencing the infamous 2019 OpenAI decision to initially withhold GPT-2-large from public release due to fears it was 'too dangerous,' ...

model-releasesyann-lecun--x
7 Apr 2026
Model Releases

Thank you to @AnthropicAI for sending FFmpeg patches

DGX agent

Thank you to @AnthropicAI for sending FFmpeg patches Introducing Project Glasswing: an urgent initiative to help secure the world’s most critical software. It’s powered by our newest frontier model, C

model-releasesboris-cherny--x
7 Apr 2026
Research

A New Type of Adversarial Examples

DGX agent

arXiv:2510.19347v2 Announce Type: replace-cross Abstract: Most machine learning models are vulnerable to adversarial examples, which poses security concerns on these models. Adversarial examples are c

researcharxiv-cs-ai
25 Aug 2026
Safety

Aligned Alone, Misaligned Together: Forecasting Adversarial Capture in LLM Agent Populations

DGX agent

arXiv:2608.22444v1 Announce Type: new Abstract: The unit of AI safety evaluation is still the individual model, yet language-model agents are increasingly deployed in interacting populations that read

safetyarxiv-cs-cl
25 Aug 2026
Research

An Interpretable Deep Learning Framework for Material Perception and Classification from Multisensory Tactile Data

DGX agent

arXiv:2608.21894v1 Announce Type: new Abstract: Human tactile perception relies on complex multisensory cues. Yet the relationship between tactile signals and perceptual representations remains poorly

researcharxiv-cs-ro
25 Aug 2026
Model Releases

Beyond Benchmarks: LLM Evaluation with an Anthropomorphic and Lifecycle-oriented Roadmap

DGX agent

arXiv:2508.18646v3 Announce Type: replace Abstract: Despite their rapid advancement, large language models (LLMs) suffer from a critical disconnect between benchmark scores and real-world utility. Cur

model-releasesarxiv-cs-ai
25 Aug 2026
Model Releases

BLADE: Bilevel Low-rank Augmented-Lagrangian Erasure for LLM Unlearning

DGX agent

arXiv:2608.22557v1 Announce Type: cross Abstract: Existing LLM unlearning methods struggle with robustness: unbounded forget losses degrade model coherence, fixed-weight balancing cannot adapt as reta

model-releasesarxiv-cs-ai
25 Aug 2026
Model Releases

CatchBench: When Can an Agent Failure Be Caught?

DGX agent

arXiv:2608.22808v1 Announce Type: new Abstract: When can an agent failure be caught? An audit is usually limited by the record rather than by the method. CatchBench therefore puts one auditor's questi

model-releasesarxiv-cs-lg
25 Aug 2026
Research

Cognitive Profiling of LRMs' Reasoning Traces Using Bloom's Taxonomy

DGX agent

arXiv:2608.23205v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have revolutionized reasoning in LLMs, and the increasing public availability of reasoning traces creates valuable opportu

researcharxiv-cs-ai
25 Aug 2026
Tools

Compute has always been core to the @VerneRobotics story. Getting there meant negotiating long, expensive contracts with big neoclouds, and …

DGX agent

Compute has always been core to the @VerneRobotics story. Getting there meant negotiating long, expensive contracts with big neoclouds, and projecting compute needs years into the future. Verne wasn't

toolstogether-ai--x
25 Aug 2026
Applications

Conformal Risk Minimization for Semi-Supervised Domain Adaptation via Optimal Transport

DGX agent

arXiv:2608.23153v1 Announce Type: new Abstract: In high-stakes healthcare applications, machine learning models are frequently trained on data from one patient population and deployed on another, crea

applicationsarxiv-cs-lg
25 Aug 2026
Model Releases

Crossing the Margin Cliff: Toward Relearn-Robust LLM Unlearning via Margin Calibration

DGX agent

arXiv:2607.27836v2 Announce Type: replace Abstract: Large language model unlearning is consistently fragile under relearn attacks. On TOFU, fine-tuning on twenty forget examples substantially recovers

model-releasesarxiv-cs-ai
25 Aug 2026
Model Releases

Cultural Moment Benchmark: Evaluating Video Cultural Reasoning and Grounding in Southeast Asia

DGX agent

arXiv:2608.23065v1 Announce Type: cross Abstract: Cultural understanding in video means more than recognizing what is visible; it requires grasping the symbolic and temporal significance of cultural c

model-releasesarxiv-cs-ai
25 Aug 2026
Model Releases

DiaRelay: Relaying Dialogue Context with a Constant-Size Memory for Emotion Recognition in Conversation

DGX agent

arXiv:2608.22745v1 Announce Type: cross Abstract: Emotion Recognition in Conversation (ERC) requires models to identify subtle emotional cues that are often distributed across distant dialogue turns.

model-releasesarxiv-cs-ai
25 Aug 2026
← Previous
1…482483484485486…1359
Next →