AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,370 results
19 Apr 2026

As AI powers Google, what’s next for Google Cloud

AgentsDGX agent

The agentic artificial intelligence era is forcing a reset in enterprise architecture. Agents that take action go well beyond analyzing data sitting in lakehouses. When agents operate on behalf of hum

b8842

Local AiDGX agent

b8842 is a release of llama.cpp, a C/C++ implementation for LLM inference. The llama.cpp project publishes multiple releases in a single day as part of its active development cycle. This specific rele

Great paper on self-improving agents. Why? We need to think more deeply about AI agent system design. The protocol specifies a framework for…

AgentsDGX agent

Great paper on self-improving agents. Why? We need to think more deeply about AI agent system design. The protocol specifies a framework for proposing, assessing, and committing improvements with audi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Sakana AIが昨年公開した日本語金融ベンチマーク「EDINET-Bench」が、国際会議 #ICLR2026 に採択されました。 ブログ:https://sakana.ai/edinet-bench/ EDINET-Benchは、金融庁EDINETの有価証券報告書約41,0…

ResearchDGX agent

Sakana AIが昨年公開した日本語金融ベンチマーク「EDINET-Bench」が、国際会議 #ICLR2026 に採択されました。 ブログ:https://sakana.ai/edinet-bench/ EDINET-Benchは、金融庁EDINETの有価証券報告書約41,000件をもとに、会計不正検知・業績予想・業種予測の3タスクでLLMを評価するベンチマークです。 公開以降、日本の金融分野

Tesla vient d'allumer à Corpus Christi (çà ne s'invente pas...) la première raffinerie majeure de lithium des États-Unis, et ce qui me frapp…

IndustryDGX agent

Tesla vient d'allumer à Corpus Christi (çà ne s'invente pas...) la première raffinerie majeure de lithium des États-Unis, et ce qui me frappe n'est pas le made in USA. C'est que Musk a balancé l'appro

Thank you @scaryrawr for the help! https://github.com/ollama/ollama/pull/15583

Local AiDGX agent

This post is a thank you acknowledgment from the Ollama project to a contributor (@scaryrawr) for their assistance with pull request #15583 on the Ollama GitHub repository. The tweet references a spec

You ever go on Huggingface and see: - GGUF - Unsloth - Llama.cpp - Dynamic GGUF - Q_4_M / IQ_4XL etc. Here's what's going on under the hood.…

Model ReleasesDGX agent

This post explains the technical details behind common terms and tools encountered on Hugging Face for running large language models locally, including quantization formats (GGUF, Q_4_M, IQ_4XL), opti

18 Apr 2026

Build your own assistant with @NVIDIAAI. ❤️

Local AiDGX agent

Build your own assistant with @NVIDIAAI. ❤️ Here's your weekend project. Build a fully local, sandboxed AI assistant. Step-by-step tutorial to build your always-on agent: 🦞 on openclaw ✅ with NVIDIA N

“my honest read is that a significant portion of this spending is driven by competitive fear rather than demonstrated returns. Nobody wants …

SafetyDGX agent

“my honest read is that a significant portion of this spending is driven by competitive fear rather than demonstrated returns. Nobody wants to be the company that didn't invest in AI when everyone els

在 Ollama 运行 hermes

Local AiDGX agent

This post likely discusses how to run the Hermes language model using Ollama, an open-source tool for running large language models locally. It probably provides instructions or insights on setting up

日経ポッドキャストにて、Sakana AIの防衛領域での取り組みを取り上げていただきました。 「防衛のAI、国産に託す 日本発サカナの分析システム受注」というテーマで、NIKKEI Digital Governanceの中西豊紀編集長に解説いただいています。ぜひお聴きください。 …

ResearchDGX agent

日経ポッドキャストにて、Sakana AIの防衛領域での取り組みを取り上げていただきました。 「防衛のAI、国産に託す 日本発サカナの分析システム受注」というテーマで、NIKKEI Digital Governanceの中西豊紀編集長に解説いただいています。ぜひお聴きください。 @nikkeipodcast https://www.nikkei.com/article/DGXZQOUC27ALB0

17 Apr 2026

china discount is real, damn

ToolsDGX agent

This post likely discusses the 'China discount' phenomenon, where Chinese companies or products are priced significantly lower than Western equivalents, examining whether this cost advantage is genuin

Chinese Language Is Not More Efficient Than English in Vibe Coding: A Preliminary Study on Token Cost and Problem-Solving Rate

Model ReleasesDGX agent

arXiv:2604.14210v1 Announce Type: new Abstract: A claim has been circulating on social media and practitioner forums that Chinese prompts are more token-efficient than English for LLM coding tasks, po

Cognitive Offloading in Agile Teams: How Artificial Intelligence Reshapes Risk Assessment and Planning Quality

ResearchDGX agent

arXiv:2604.13814v1 Announce Type: cross Abstract: Recent advances in artificial intelligence (AI) have shown promise in automating key aspects of Agile project management, yet their impact on team cog

Conformal Policy Control

SafetyDGX agent

arXiv:2603.02196v2 Announce Type: replace-cross Abstract: An agent must try new behaviors to explore and improve. In high-stakes environments, an agent that violates safety constraints may cause harm

cool new paper on self-improving agents

AgentsDGX agent

cool new paper on self-improving agents // Self-Evolving Agent Protocol // One of the more interesting papers I read this week. (bookmark it if you are an AI dev) The paper introduces Autogenesis, a s

Create Expert Content: Deploying a Multi-Agent System with Terraform and Cloud Run

Model ReleasesDGX agent

In support of our mission to accelerate the developer journey on Google Cloud, we built Dev Signal: a multi-agent system designed to transform raw community signals into reliable technical guidance by

Data Driven Agent Design with Evals & Hill Climbing Algorithms this is a mental model dump i’ve been thinking through + iterating on as we’r…

AgentsDGX agent

Data Driven Agent Design with Evals & Hill Climbing Algorithms this is a mental model dump i’ve been thinking through + iterating on as we’re building self-improvement infra around agents: - mining Tr

DPSQL+: A Differentially Private SQL Library with a Minimum Frequency Rule

Model ReleasesDGX agent

arXiv:2602.22699v2 Announce Type: replace-cross Abstract: SQL is the de facto interface for exploratory data analysis; however, releasing exact query results can expose sensitive information through m

EchoAgent: Towards Reliable Echocardiography Interpretation with 'Eyes','Hands' and 'Minds'

AgentsDGX agent

arXiv:2604.05541v2 Announce Type: replace Abstract: Reliable interpretation of echocardiography (Echo) is crucial for assessing cardiac function, which demands clinicians to synchronously orchestrate

Enhancing LLM-Based Neural Network Generation: Few-Shot Prompting and Efficient Validation for Automated Architecture Design

ResearchDGX agent

arXiv:2512.24120v2 Announce Type: replace Abstract: Automated neural network architecture design remains a significant challenge in computer vision. Task diversity and computational constraints requir

Exploration and Exploitation Errors Are Measurable for Language Model Agents

SafetyDGX agent

arXiv:2604.13151v1 Announce Type: new Abstract: Language Model (LM) agents are increasingly used in complex open-ended decision-making tasks, from AI coding to physical AI. A core requirement in these

FRAGATA: Semantic Retrieval of HPC Support Tickets via Hybrid RAG over 20 Years of Request Tracker History

TutorialsDGX agent

arXiv:2604.13721v1 Announce Type: cross Abstract: The technical support team of a supercomputing centre accumulates, over the course of decades, a large volume of resolved incidents that constitute cr

From Black Box to Glass Box: Cross-Model ASR Disagreement to Prioto Review in Ambient AI Scribe Documentation

ApplicationsDGX agent

arXiv:2604.14152v1 Announce Type: cross Abstract: Ambient AI 'scribe' systems promise to reduce clinical documentation burden, but automatic speech recognition (ASR) errors can remain unnoticed withou

From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability

HardwareDGX agent

arXiv:2604.13048v1 Announce Type: cross Abstract: Modern cloud-native platforms expose thousands of time series metrics through systems like Prometheus, yet formulating correct queries in domain-speci

FWIW i have just updated chrome and do not see this. google has this big issue of ultra slow incremental rollouts. it really kills the vibe …

ToolsDGX agent

FWIW i have just updated chrome and do not see this. google has this big issue of ultra slow incremental rollouts. it really kills the vibe - i see stuff they launch, i am excited to try out, 'oh its

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds

ResearchDGX agent

arXiv:2604.14268v1 Announce Type: new Abstract: We introduce HY-World 2.0, a multi-modal world model framework that advances our prior project HY-World 1.0. HY-World 2.0 accommodates diverse input mod

I don't understand why we automatically give people credit for 'sincere views.' Like who gives a fuck? If I have a sincere view that a trans…

ResearchDGX agent

I don't understand why we automatically give people credit for 'sincere views.' Like who gives a fuck? If I have a sincere view that a transdimensional vampire attack is imminent it doesn't make it sa

in retrospect putting the slop cannons (@_lopopolo) on @aiDotEngineer talks day 1 and putting the grown ups (@badlogicgames) on talks day 2 …

ToolsDGX agent

in retrospect putting the slop cannons (@_lopopolo) on @aiDotEngineer talks day 1 and putting the grown ups (@badlogicgames) on talks day 2 is working out pretty well for faithfully representing the m

Is there still a widespread belief that LLMs and coding agents are good for greenfield development but don't help for maintaining large exis…

ToolsDGX agent

Simon Willison discusses the perception that LLMs and coding agents are primarily useful for greenfield development (starting new projects from scratch) rather than for maintaining and modifying large

Isn’t just what to remember, but when to update it: in the loop or after the fact. A must read by @hwchase17!

AgentsDGX agent

This post by Harrison Chase (LangChain creator) discusses the importance of timing in updating information systems or memory mechanisms in AI applications, comparing real-time updates ('in the loop')

Join us at PyCon US 2026 in Long Beach - we have new AI and security tracks this year

Model ReleasesDGX agent

This year's PyCon US is coming up next month from May 13th to May 19th, with the core conference talks from Friday 15th to Sunday 17th and tutorial and sprint days either side. It's in Long Beach, Cal

Label-efficient underwater species classification with logistic regression on frozen foundation model embeddings

Model ReleasesDGX agent

arXiv:2604.00313v2 Announce Type: replace Abstract: Automated species classification from underwater imagery is bottlenecked by the cost of expert annotation, and supervised models trained on one data

Language Model Fine-Tuning on Scaled Survey Data for Predicting Distributions of Public Opinions

ApplicationsDGX agent

arXiv:2502.16761v2 Announce Type: replace Abstract: Large language models (LLMs) present novel opportunities in public opinion research by predicting survey responses in advance during the early stage

Large Language Models to Enhance Business Process Modeling: Past, Present, and Future Trends

ApplicationsDGX agent

arXiv:2604.14034v1 Announce Type: cross Abstract: Recent advances in Generative Artificial Intelligence, particularly Large Language Models (LLMs), have stimulated growing interest in automating or as

Last week, Anthropic announced Project Glasswing alongside Claude Mythos Preview, a model they described as so powerful at finding vulnerabi…

Model ReleasesDGX agent

Last week, Anthropic announced Project Glasswing alongside Claude Mythos Preview, a model they described as so powerful at finding vulnerabilities they couldn't release it. The announcement featured A

LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.14922v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a critical driver for enhancing the reasoning capabilities of Large Language Models (LLMs). While recent ad

@_lopopolo @aiDotEngineer @badlogicgames Slow Down talk: https://x.com/aiDotEngineer/status/2044915752072036420?s=20

ToolsDGX agent

@_lopopolo @aiDotEngineer @badlogicgames Slow Down talk: https://x.com/aiDotEngineer/status/2044915752072036420?s=20 🆕 Building pi in a World of Slop https://www.youtube.com/watch?v=RjfbvDXpFls @badlo

Low-Cost System for Automatic Recognition of Driving Pattern in Assessing Interurban Mobility using Geo-Information

ResearchDGX agent

arXiv:2604.15216v1 Announce Type: cross Abstract: Mobility in urban and interurban areas, mainly by cars, is a day-to-day activity of many people. However, some of its main drawbacks are traffic jams

MCPThreatHive: Automated Threat Intelligence for Model Context Protocol Ecosystems

AgentsDGX agent

arXiv:2604.13849v1 Announce Type: cross Abstract: The rapid proliferation of Model Context Protocol (MCP)-based agentic systems has introduced a new category of security threats that existing framewor

MedVerse: Efficient and Reliable Medical Reasoning via DAG-Structured Parallel Execution

ResearchDGX agent

arXiv:2602.07529v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated strong performance and rapid progress in a wide range of medical reasoning tasks. However, their sequ

Model Capability Dominates: Inference-Time Optimization Lessons from AIMO 3

HardwareDGX agent

arXiv:2603.27844v2 Announce Type: replace Abstract: Majority voting over multiple LLM attempts improves mathematical reasoning, but correlated errors limit the effective sample size. A natural fix is

Polyformer: a generative framework for thermodynamic modeling of polymeric molecules

ResearchDGX agent

arXiv:2604.14241v1 Announce Type: cross Abstract: The classic paradigm of structural biology is that the sequence of a biomolecule (protein, nucleic acid, lipid, etc) determines its conformation (shap

Psychological Steering of Large Language Models

ResearchDGX agent

arXiv:2604.14463v1 Announce Type: new Abstract: Large language models (LLMs) emulate a consistent human-like behavior that can be shaped through activation-level interventions. This paradigm is conver

R3D: Revisiting 3D Policy Learning

SafetyDGX agent

arXiv:2604.15281v1 Announce Type: new Abstract: 3D policy learning promises superior generalization and cross-embodiment transfer, but progress has been hindered by training instabilities and severe o

Sources: Cursor is in advanced talks to raise about 2B co-led by a16z at a pre-money valuation of more than 50B, with Nvidia participating (Bloomberg)

HardwareDGX agent

Bloomberg: Sources: Cursor is in advanced talks to raise about 2B co-led by a16z at a pre-money valuation of more than 50B, with Nvidia participating — Cursor, a leading artificial intelligence startu

The Code Whisperer: LLM and Graph-Based AI for Smell and Vulnerability Resolution

ResearchDGX agent

arXiv:2604.13114v1 Announce Type: cross Abstract: Code smells and software vulnerabilities both increase maintenance cost, yet they are often handled by separate tools that miss structural context and

'The post all LLMs hate and want you to NOT SEE' ^^ Serious title, you can also discuss this with the LLM, it would be biased against it, as…

AgentsDGX agent

'The post all LLMs hate and want you to NOT SEE' ^^ Serious title, you can also discuss this with the LLM, it would be biased against it, as it's known at this point that memory is the true long term

The Specification Trap: Why Static Value Alignment Alone Is Insufficient for Robust Alignment

SafetyDGX agent

arXiv:2512.03048v4 Announce Type: replace-cross Abstract: Static content-based AI value alignment is insufficient for robust alignment under capability scaling, distributional shift, and increasing au

Towards Trustworthy 6G Network Digital Twins: A Framework for Validating Counterfactual What-If Analysis in Edge Computing Resources

SafetyDGX agent

arXiv:2604.14787v1 Announce Type: cross Abstract: Network Digital Twins (NDTs) enable safe what-if analysis for 6G cloud-edge infrastructures, but adoption is often limited by fragmented workflows fro

Which bird does not have wings: Negative-constrained KGQA with Schema-guided Semantic Matching and Self-directed Refinement

ApplicationsDGX agent

arXiv:2604.14749v1 Announce Type: new Abstract: Large language models still struggle with faithfulness and hallucinations despite their remarkable reasoning abilities. In Knowledge Graph Question Answ

WybeCoder: Verified Imperative Code Generation

AgentsDGX agent

arXiv:2603.29088v2 Announce Type: replace-cross Abstract: Recent progress in large language models (LLMs) has substantially advanced automatic code generation and formal theorem proving, yet software

16 Apr 2026

81% of Grade 4 children in South Africa (that’s 914,000 kids out of 1.1 million) cannot read for MEANING in ANY of South Africa’s 11 languag…

TutorialsDGX agent

81% of Grade 4 children in South Africa (that’s 914,000 kids out of 1.1 million) cannot read for MEANING in ANY of South Africa’s 11 languages! It got WORSE — up from 78% in 2016. South Africa ranked

A Complete Symmetry Classification of Shallow ReLU Networks

Model ReleasesDGX agent

arXiv:2604.14037v1 Announce Type: new Abstract: Parameter space is not function space for neural network architectures. This fact, investigated as early as the 1990s under terms such as ``reverse engi

Adithya is putting his work where his heart is. If you believe in open research and open AI, join a company that actually lives those values…

IndustryDGX agent

Adithya is putting his work where his heart is. If you believe in open research and open AI, join a company that actually lives those values, not a closed-source, revenue-maximizing one! Quick career

AI for Materials Science starter kit [D]

ResearchDGX agent

This r/MachineLearning discussion post serves as a community-curated beginner's resource for applying artificial intelligence and machine learning to materials science, likely compiling recommended to

Atelier: a canvas for thinking and making with local models.

Local AiDGX agent

Atelier is a canvas-like system that leverages generative image and video models to blend spaces for thinking and creation, where both references and generated assets co-exist in one unified workspace

b8808

Local AiDGX agent

Build b8808 is an incremental release of **llama.cpp**, the open-source C/C++ library for running large language model (LLM) inference locally. Like all llama.cpp builds, it likely includes bug fixes,

b8814

Local AiDGX agent

llama.cpp is a C/C++ library for LLM inference that enables running large language models on consumer hardware. Release b8814 is a specific version in the project's continuous release cycle, which fol

BOAT: Navigating the Sea of In Silico Predictors for Antibody Design via Multi-Objective Bayesian Optimization

ResearchDGX agent

arXiv:2604.13980v1 Announce Type: new Abstract: Antibody lead optimization is inherently a multi-objective challenge in drug discovery. Achieving a balance between different drug-like properties is cr

← Previous
1…8384858687…90
Next →