AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,370 results
21 Apr 2026

Late Fusion Neural Operators for Extrapolation Across Parameter Space in Partial Differential Equations

Model ReleasesDGX agent

arXiv:2604.16721v1 Announce Type: new Abstract: Developing neural operators that accurately predict the behavior of systems governed by partial differential equations (PDEs) across unseen parameter re

Lightweight Cybersickness Detection based on User-Specific Eye and Head Tracking Data in Virtual Reality

ApplicationsDGX agent

arXiv:2604.17158v1 Announce Type: cross Abstract: The occurrence of cybersickness in virtual reality (VR) significantly impairs users' perception and sense of immersion. Therefore, timely detection of

Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies

Safety

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.16516v1 Announce Type: new Abstract: Text-to-Image (T2I) generation models have been widely adopted across various industries, yet are criticized for frequently exhibiting societal stereoty

Pearmut: Human Evaluation of Translation Made Trivial

ResearchDGX agent

arXiv:2601.02933v3 Announce Type: replace Abstract: Human evaluation is the gold standard for multilingual NLP, but is often skipped in practice and substituted with automatic metrics because it is no

Physics-Informed Neural Networks: A Didactic Derivation of the Complete Training Cycle

Model ReleasesDGX agent

arXiv:2604.18481v1 Announce Type: cross Abstract: This paper is a step-by-step, self-contained guide to the complete training cycle of a Physics-Informed Neural Network (PINN) -- a topic that existing

ProTrain: Efficient LLM Training via Memory-Aware Techniques

Model ReleasesDGX agent

arXiv:2406.08334v2 Announce Type: replace-cross Abstract: Memory pressure has emerged as a dominant constraint in scaling the training of large language models (LLMs), particularly in resource-constra

RAG-DIVE: A Dynamic Approach for Multi-Turn Dialogue Evaluation in Retrieval-Augmented Generation

ApplicationsDGX agent

arXiv:2604.16310v1 Announce Type: cross Abstract: Evaluating Retrieval-Augmented Generation (RAG) systems using static multi-turn datasets fails to capture the dynamic nature of real-world dialogues.

ReasoningBank: Enabling agents to learn from experience

TutorialsDGX agent

ReasoningBank is a memory framework that distills generalizable reasoning strategies from an agent's successful and failed experiences, which the agent retrieves at test time to inform interactions wh

ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2603.05863v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have revolutionized code generation, standard ``System 1'' approaches that generate solutions in a single forward

Scaling Human-AI Coding Collaboration Requires a Governable Consensus Layer

Model ReleasesDGX agent

arXiv:2604.17883v1 Announce Type: cross Abstract: Vibe coding produces correct, executable code at speed, but leaves no record of the structural commitments, dependencies, or evidence behind it. Revie

SciDraw-6K: A Multilingual Scientific Illustration Dataset Generated by Google Gemini

Model ReleasesDGX agent

arXiv:2604.17206v1 Announce Type: new Abstract: We present SciDraw-6K, a curated dataset of 6,291 scientific illustrations synthesized by Google Gemini image-generation models, each paired with prompt

SpatialImaginer: Towards Adaptive Visual Imagination for Spatial Reasoning

ResearchDGX agent

arXiv:2604.17385v1 Announce Type: new Abstract: Spatial intelligence, which refers to the ability to reason about geometric and physical structure from visual observations, remains a core challenge fo

SynthPID: P&ID digitization from Topology-Preserving Synthetic Data

Model ReleasesDGX agent

arXiv:2604.16513v1 Announce Type: new Abstract: Automating the digitization of Piping and Instrumentation Diagrams (P&IDs) into structured process graphs would unlock significant value in plant operat

Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks

Model ReleasesDGX agent

arXiv:2604.17159v1 Announce Type: cross Abstract: We present, to our knowledge, the most comprehensive cross-model evaluation of LLM agents on offensive cybersecurity tasks, benchmarking 10 frontier m

Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards

AgentsDGX agent

arXiv:2603.02008v2 Announce Type: replace Abstract: Effective exploration in reinforcement learning requires not only tracking where an agent has been, but also understanding how the agent perceives a

Textual Bayes: Quantifying Prompt Uncertainty in LLM-Based Systems

ApplicationsDGX agent

arXiv:2506.10060v2 Announce Type: replace Abstract: Although large language models (LLMs) are becoming increasingly capable of solving challenging real-world tasks, accurately quantifying their uncert

TGLF-WINN: Data-Efficient Deep Learning Surrogate for Turbulent Transport Modeling in Fusion

ResearchDGX agent

arXiv:2509.07024v2 Announce Type: replace-cross Abstract: The Trapped Gyro-Landau Fluid (TGLF) model provides fast, accurate predictions of turbulent transport in tokamaks, but whole device simulation

The Cognitive Penalty: Ablating System 1 and System 2 Reasoning in Edge-Native SLMs for Decentralized Consensus

Model ReleasesDGX agent

arXiv:2604.16913v1 Announce Type: cross Abstract: Decentralized Autonomous Organizations (DAOs) are inclined explore Small Language Models (SLMs) as edge-native constitutional firewalls to vet proposa

Time Series Forecasting as Reasoning: A Slow-Thinking Approach with Reinforced LLMs

SafetyDGX agent

arXiv:2506.10630v2 Announce Type: replace Abstract: To advance time series forecasting (TSF), various methods have been proposed to improve prediction accuracy, evolving from statistical techniques to

Towards Reliable Testing of Machine Unlearning

ResearchDGX agent

arXiv:2604.16536v1 Announce Type: new Abstract: Machine learning components are now central to AI-infused software systems, from recommendations and code assistants to clinical decision support. As re

Unraveling the Key of Machine Learning-based Android Malware Detection

TutorialsDGX agent

arXiv:2402.02953v2 Announce Type: replace-cross Abstract: With the rapid advancement of machine learning (ML), ML-based Android malware detection has gained significant popularity due to its ability t

Upper Approximation Bounds for Neural Oscillators

ResearchDGX agent

arXiv:2512.01015v2 Announce Type: replace Abstract: Neural oscillators, originating from second-order ordinary differential equations (ODEs), have demonstrated strong performance in stably learning ca

We're honored to be named Google Cloud's 2026 AI Tooling Partner of the Year. This recognition reflects what 50 million builders have made p…

ApplicationsDGX agent

We're honored to be named Google Cloud's 2026 AI Tooling Partner of the Year. This recognition reflects what 50 million builders have made possible together. Product managers, founders, students, oper

What's the deal with spacesuits for the Moon? Will they be ready in time?

IndustryDGX agent

NASA's next-generation spacesuit for the Artemis III mission, the AxEMU developed by Axiom Space, has passed a contractor-led technical review and continues undergoing testing with NASA astronauts and

20 Apr 2026

A Comparative Study on the Impact of Traditional Learning and Interactive Learning on Students' Academic Performance and Emotional Well-Being

TutorialsDGX agent

arXiv:2604.15335v1 Announce Type: cross Abstract: The growing adoption of interactive learning tools in higher education offers new opportunities to enhance student performance and well-being. This st

An Information-Geometric Approach to Artificial Curiosity

Model ReleasesDGX agent

arXiv:2504.06355v2 Announce Type: replace Abstract: Learning in environments with sparse rewards remains a fundamental challenge in reinforcement learning. Artificial curiosity addresses this limitati

Analyzing Chain of Thought (CoT) Approaches in Control Flow Code Deobfuscation Tasks

TutorialsDGX agent

arXiv:2604.15390v1 Announce Type: cross Abstract: Code deobfuscation is the task of recovering a readable version of a program while preserving its original behavior. In practice, this often requires

and @dexhorthy is quoting Z/L continuum in AIE Miami!! https://x.com/altryne/status/2043748676099866771?s=46 idea catching on @altryne

ToolsDGX agent

This post references @dexhorthy quoting the Z/L continuum concept at an AIE Miami event, suggesting the idea is gaining traction in AI discussions. The Z/L continuum appears to be a framework or model

Applied Explainability for Large Language Models: A Comparative Study

ApplicationsDGX agent

arXiv:2604.15371v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance across many natural language processing tasks, yet their decision processes remain difficult t

AutoFed: Personalized Federated Traffic Prediction via Adaptive Prompt

Model ReleasesDGX agent

arXiv:2512.24625v2 Announce Type: replace-cross Abstract: Accurate traffic prediction is essential for Intelligent Transportation Systems, including ride-hailing, urban road planning, and vehicle flee

Automating Crash Diagram Generation Using Vision-Language Models: A Case Study on Multi-Lane Roundabouts

Model ReleasesDGX agent

arXiv:2604.15332v1 Announce Type: cross Abstract: Crash diagrams are essential tools in transportation safety analysis, yet their manual preparation remains time-consuming and prone to human variabili

b8853

Local AiDGX agent

b8853 is a release of llama.cpp , the open-source C/C++ library for running large language model inference. The project enables LLM inference with minimal setup and state-of-the-art performance on a w

Bilevel Optimization of Agent Skills via Monte Carlo Tree Search

AgentsDGX agent

arXiv:2604.15709v1 Announce Type: new Abstract: Agent exttt{skills} are structured collections of instructions, tools, and supporting resources that help large language model (LLM) agents perform part

Capture the Flags: Family-Based Evaluation of Agentic LLMs via Semantics-Preserving Transformations

AgentsDGX agent

arXiv:2602.05523v2 Announce Type: replace-cross Abstract: Agentic large language models (LLMs) are increasingly evaluated on cybersecurity tasks using capture-the-flag (CTF) benchmarks, yet existing p

ChemAmp: Amplified Chemistry Tools via Composable Agents

AgentsDGX agent

arXiv:2505.21569v3 Announce Type: replace-cross Abstract: Although LLM-based agents are proven to master tool orchestration in scientific fields, particularly chemistry, their single-task performance

Claude Design rocks and Claude Design sucks. Let's talk about. Last night, I created an entire game in Claude Design inspired by the Longyou…

Model ReleasesDGX agent

Claude Design rocks and Claude Design sucks. Let's talk about. Last night, I created an entire game in Claude Design inspired by the Longyou Caves in China (why is there no soot!), and we all gathered

dear god lol - new kimi model is a fucking beast. GPT 5.4 level coding, 76% cheaper than opus 4.7 and 100% open source / free to use, i mean…

Model ReleasesDGX agent

dear god lol - new kimi model is a fucking beast. GPT 5.4 level coding, 76% cheaper than opus 4.7 and 100% open source / free to use, i mean look at this: > kimi k2.6 can code continuously for 12 hour

Efficient Video Diffusion Models: Advancements and Challenges

ApplicationsDGX agent

arXiv:2604.15911v1 Announce Type: new Abstract: Video diffusion models have rapidly become the dominant paradigm for high-fidelity generative video synthesis, but their practical deployment remains co

EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems

Model ReleasesDGX agent

arXiv:2510.13220v2 Announce Type: replace Abstract: A fundamental limitation of current AI agents is their inability to learn complex skills on the fly at test time, often behaving like 'clever but cl

Explainable Iterative Data Visualisation Refinement via an LLM Agent

AgentsDGX agent

arXiv:2604.15319v1 Announce Type: cross Abstract: Exploratory analysis of high-dimensional data relies on embedding the data into a low-dimensional space (typically 2D or 3D), based on which visualiza

From Multi-Agent to Single-Agent: When Is Skill Distillation Beneficial?

AgentsDGX agent

arXiv:2604.01608v2 Announce Type: replace Abstract: Multi-agent systems (MAS) tackle complex tasks by distributing expertise, though this often comes at the cost of heavy coordination overhead, contex

From Seeing to Simulating: Generative High-Fidelity Simulation with Digital Cousins for Generalizable Robot Learning and Evaluation

ApplicationsDGX agent

arXiv:2604.15805v1 Announce Type: cross Abstract: Learning robust robot policies in real-world environments requires diverse data augmentation, yet scaling real-world data collection is costly due to

GIST: Multimodal Knowledge Extraction and Spatial Grounding via Intelligent Semantic Topology

Local AiDGX agent

arXiv:2604.15495v1 Announce Type: new Abstract: Navigating complex, densely packed environments like retail stores, warehouses, and hospitals poses a significant spatial grounding challenge for humans

how to deploy long horizon agents, and all the infra you need!

TutorialsDGX agent

This post likely discusses the infrastructure and practical considerations required for deploying AI agents capable of handling long-horizon tasks—those requiring multiple steps and extended planning

InfoChess: A Game of Adversarial Inference and a Laboratory for Quantifiable Information Control

AgentsDGX agent

arXiv:2604.15373v1 Announce Type: cross Abstract: We propose InfoChess, a symmetric adversarial game that elevates competitive information acquisition to the primary objective. There is no piece captu

Intelligent Healthcare Imaging Platform: A VLM-Based Framework for Automated Medical Image Analysis and Clinical Report Generation

Model ReleasesDGX agent

arXiv:2509.13590v3 Announce Type: replace-cross Abstract: The rapid advancement of artificial intelligence (AI) in healthcare imaging has revolutionized diagnostic medicine and clinical decision-makin

@Kimi_Moonshot Kimi-K2.6 just dropped in Qoder. SOTA coding · Long-horizon execution · Agent swarms At 0.3x credits.

AgentsDGX agent

Kimi-K2.6 is a new AI model released through Qoder that achieves state-of-the-art performance in coding tasks and supports long-horizon execution and agent swarms, available at a reduced cost of 0.3x

LLM attribution analysis across different fine-tuning strategies and model scales for automated code compliance

Model ReleasesDGX agent

arXiv:2604.15589v1 Announce Type: cross Abstract: Existing research on large language models (LLMs) for automated code compliance has primarily focused on performance, treating the models as black box

LLMbench: A Comparative Close Reading Workbench for Large Language Models

ResearchDGX agent

arXiv:2604.15508v1 Announce Type: cross Abstract: LLMbench is a browser-based workbench for the comparative close reading of large language model (LLM) outputs. Where existing tools for LLM comparison

MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report Generation

AgentsDGX agent

arXiv:2604.16175v1 Announce Type: new Abstract: Automated 3D radiology report generation often suffers from clinical hallucinations and a lack of the iterative verification found in human practice. Wh

Model page https://ollama.com/library/kimi-k2.6 More integrations https://docs.ollama.com/integrations

Local AiDGX agent

Ollama has made the Kimi K2.6 model available in its library, allowing users to run this model locally through the Ollama platform. The announcement highlights expanded integration options documented

PIIBench: A Unified Multi-Source Benchmark Corpus for Personally Identifiable Information Detection

Model ReleasesDGX agent

arXiv:2604.15776v1 Announce Type: cross Abstract: We present PIIBench, a unified benchmark corpus for Personally Identifiable Information (PII) detection in natural language text. Existing resources f

RaveDAO's RAVE has lost $6.6B+ in market cap and its price has sunk ~98% since Saturday, after ZachXBT called on exchanges to probe if it was being manipulated (André Beganski/Decrypt)

IndustryDGX agent

André Beganski / Decrypt: RaveDAO's RAVE has lost $6.6B+ in market cap and its price has sunk ~98% since Saturday, after ZachXBT called on exchanges to probe if it was being manipulated — RaveDAO's to

Rogue Trooper brings the Genetic Infantry to the silver screen

IndustryDGX agent

Rogue Trooper is an upcoming adult animated military science fiction film based on the 2000 AD comic strip, written and directed by Duncan Jones. The film follows 19, a 'Genetic Infantryman' who is th

The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination

Model ReleasesDGX agent

arXiv:2510.22977v2 Announce Type: replace-cross Abstract: Enhancing the reasoning capabilities of Large Language Models (LLMs) is a key strategy for building Agents that 'think then act.' However, rec

The Relic Condition: When Published Scholarship Becomes Material for Its Own Replacement

Model ReleasesDGX agent

arXiv:2604.16116v1 Announce Type: cross Abstract: We extracted the scholarly reasoning systems of two internationally prominent humanities and social science scholars from their published corpora alon

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning

AgentsDGX agent

arXiv:2411.10446v3 Announce Type: replace-cross Abstract: Recent progress in vision-language models (VLMs) has opened new possibilities for robot task planning, but these models often produce incorrec

VeriMoA: A Mixture-of-Agents Framework for Spec-to-HDL Generation

AgentsDGX agent

arXiv:2510.27617v2 Announce Type: replace Abstract: Automation of Register Transfer Level (RTL) design can help developers meet increasing computational demands. Large Language Models (LLMs) show prom

vla-eval: A Unified Evaluation Harness for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2603.13966v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are increasingly evaluated across multiple simulation benchmarks, yet adding each benchmark to an evaluation pip

19 Apr 2026

A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M in 2026 (Emily Shugerman/The San Francisco ...)

HardwareDGX agent

Emily Shugerman / The San Francisco Standard: A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M i

← Previous
1…8283848586…90
Next →