AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
5,370 results
10 Apr 2026

AsyncTLS: Efficient Generative LLM Inference with Asynchronous Two-level Sparse Attention

ResearchDGX agent

arXiv:2604.07815v1 Announce Type: new Abstract: Long-context inference in LLMs faces the dual challenges of quadratic attention complexity and prohibitive KV cache memory. While token-level sparse att

Bird-Inspired Spatial Flapping Wing Mechanism via Coupled Linkages with Single Actuator

ResearchDGX agent

arXiv:2604.07677v1 Announce Type: new Abstract: Spatial single-loop mechanisms such as Bennett linkages offer a unique combination of one-degree-of-freedom actuation and nontrivial spatial trajectorie

Blending Human and LLM Expertise to Detect Hallucinations and Omissions in Mental Health Chatbot Responses

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.06216v1 Announce Type: cross Abstract: As LLM-powered chatbots are increasingly deployed in mental health services, detecting hallucinations and omissions has become critical for user safet

ChangeBridge: Spatiotemporal Image Generation with Multimodal Controls for Remote Sensing

ResearchDGX agent

arXiv:2507.04678v3 Announce Type: replace Abstract: Spatiotemporal image generation is a highly meaningful task, which can generate future scenes conditioned on given observations. However, existing c

ChemVLR: Prioritizing Reasoning in Perception for Chemical Vision-Language Understanding

ApplicationsDGX agent

arXiv:2604.06685v1 Announce Type: cross Abstract: While Vision-Language Models (VLMs) have demonstrated significant potential in chemical visual understanding, current models are predominantly optimiz

CubeGraph: Efficient Retrieval-Augmented Generation for Spatial and Temporal Data

ApplicationsDGX agent

arXiv:2604.06616v1 Announce Type: cross Abstract: Hybrid queries combining high-dimensional vector similarity search with spatio-temporal filters are increasingly critical for modern retrieval-augment

Digital Skin, Digital Bias: Uncovering Tone-Based Biases in LLMs and Emoji Embeddings

Model ReleasesDGX agent

arXiv:2604.06863v1 Announce Type: cross Abstract: Skin-toned emojis are crucial for fostering personal identity and social inclusion in online communication. As AI models, particularly Large Language

DosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AI

Model ReleasesDGX agent

arXiv:2604.06280v1 Announce Type: cross Abstract: Purpose: To develop and evaluate DosimeTron, an agentic AI system for automated patient-specific MC internal radiation dosimetry in PET/CT examination

GitHub Copilot CLI for Beginners: Getting started with GitHub Copilot CLI

TutorialsDGX agent

GitHub for Beginners: Getting started with the GitHub Copilot CLI, a step-by-step tutorial. The post GitHub Copilot CLI for Beginners: Getting started with GitHub Copilot CLI appeared first on The Git

GLM 5.1 on AI Gateway

ToolsDGX agent

GLM 5.1 from Z.ai is now available on Vercel's AI Gateway with no markup and no separate provider account required. Designed for long-horizon autonomous tasks, it handles planning, execution, testi...

Grok 4.20 hitting 83% on non-hallucination. Values truth. Claude ~74%. Others sitting in the 60s… or way lower. Less guessing. More honesty …

Model ReleasesDGX agent

Grok 4.20 hitting 83% on non-hallucination. Values truth. Claude ~74%. Others sitting in the 60s… or way lower. Less guessing. More honesty when it doesn’t know. That’s a different kind of intelligenc

“I spent 1 year with a Tesla and realized one thing: I was 100% confident in my misinformed opinions. I used to be an EV skeptic. I was wron…

IndustryDGX agent

“I spent 1 year with a Tesla and realized one thing: I was 100% confident in my misinformed opinions. I used to be an EV skeptic. I was wrong. Here are the 5 things that changed my mind: Charging - I

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for yo…

Model ReleasesDGX agent

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be h

Lang2Act: Fine-Grained Visual Reasoning through Self-Emergent Linguistic Toolchains

ResearchDGX agent

arXiv:2602.13235v2 Announce Type: replace-cross Abstract: Visual Retrieval-Augmented Generation (VRAG) enhances Vision-Language Models (VLMs) by incorporating external visual documents to address a gi

Learning Without Losing Identity: Capability Evolution for Embodied Agents

SafetyDGX agent

arXiv:2604.07799v1 Announce Type: new Abstract: Embodied agents are expected to operate persistently in dynamic physical environments, continuously acquiring new capabilities over time. Existing appro

LLM-based Schema-Guided Extraction and Validation of Missing-Person Intelligence from Heterogeneous Data Sources

SafetyDGX agent

arXiv:2604.06571v1 Announce Type: cross Abstract: Missing-person and child-safety investigations rely on heterogeneous case documents, including structured forms, bulletin-style posters, and narrative

LLM Spirals of Delusion: A Benchmarking Audit Study of AI Chatbot Interfaces

Model ReleasesDGX agent

arXiv:2604.06188v1 Announce Type: cross Abstract: People increasingly hold sustained, open-ended conversations with large language models (LLMs). Public reports and early studies suggest that, in such

LNN-PINN: A Unified Physics-Only Training Framework with Liquid Residual Blocks

Model ReleasesDGX agent

arXiv:2508.08935v4 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) have attracted considerable attention for their ability to integrate partial differential equation priors i

Lost in Cultural Translation: Do LLMs Struggle with Math Across Cultural Contexts?

Model ReleasesDGX agent

arXiv:2503.18018v2 Announce Type: replace Abstract: We demonstrate that large language models' (LLMs) mathematical reasoning is culturally sensitive: testing 14 models from Anthropic, OpenAI, Google,

Love this new podcast format ❤️ super insightful!

ApplicationsDGX agent

Love this new podcast format ❤️ super insightful! 🎙️Introducing Max Agency Max Agency is a new podcast where we go deep on how the best agents are actually being built: architecture decisions, tradeof

LPM 1.0: Video-based Character Performance Model

Model ReleasesDGX agent

arXiv:2604.07823v1 Announce Type: new Abstract: Performance, the externalization of intent, emotion, and personality through visual, vocal, and temporal behavior, is what makes a character alive. Lear

Max volume for Max Agency 📣🗣️

ApplicationsDGX agent

Max volume for Max Agency 📣🗣️ 🎙️Introducing Max Agency Max Agency is a new podcast where we go deep on how the best agents are actually being built: architecture decisions, tradeoffs, evals, and every

Memory Scaling for AI Agents

IndustryDGX agent

Databricks Research introduced **MemAlign**, a memory framework for AI agents that stores past interactions as episodic memories and uses an LLM to distill them into generalized semantic rules, whi...

MiniMax M2.7 is live on AI Gateway

ToolsDGX agent

MiniMax M2.7 is now available on Vercel AI Gateway in two variants: standard and high-speed, accessible without requiring separate provider accounts. It represents a significant improvement over pr...

New Guide: Incorporating human judgment in the agent improvement loop Building agents is hard. Everyone talks about the code. What gets less…

AgentsDGX agent

New Guide: Incorporating human judgment in the agent improvement loop Building agents is hard. Everyone talks about the code. What gets less attention is how to capture domain expert knowledge and act

On Emotion-Sensitive Decision Making of Small Language Model Agents

Model ReleasesDGX agent

arXiv:2604.06562v1 Announce Type: new Abstract: Small language models (SLM) are increasingly used as interactive decision-making agents, yet most decision-oriented evaluations ignore emotion as a caus

Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization

Model ReleasesDGX agent

arXiv:2604.07343v1 Announce Type: cross Abstract: Pluralistic alignment has emerged as a critical frontier in the development of Large Language Models (LLMs), with reward models (RMs) serving as a cen

Personalizing ChatGPT

TutorialsDGX agent

The OpenAI Academy 'Personalizing ChatGPT' tutorial covers several methods users can employ to tailor ChatGPT's behavior to their needs, including custom instructions, memory, and prompt engineerin...

Physics-Informed Functional Link Constrained Framework with Domain Mapping for Solving Bending Analysis of an Exponentially Loaded Perforated Beam

Model ReleasesDGX agent

arXiv:2604.07025v1 Announce Type: cross Abstract: This article presents a novel and comprehensive approach for analyzing bending behavior of the tapered perforated beam under an exponential load. The

Production-Ready Automated ECU Calibration using Residual Reinforcement Learning

ApplicationsDGX agent

arXiv:2604.07059v1 Announce Type: new Abstract: Electronic Control Units (ECUs) have played a pivotal role in transforming motorcars of yore into the modern vehicles we see on our roads today. They ac

Qualixar OS: A Universal Operating System for AI Agent Orchestration

SafetyDGX agent

arXiv:2604.06392v1 Announce Type: new Abstract: We present Qualixar OS, the first application-layer operating system for universal AI agent orchestration. Unlike kernel-level approaches (AIOS) or sing

Reasoning Graphs: Deterministic Agent Accuracy through Evidence-Centric Chain-of-Thought Feedback

AgentsDGX agent

arXiv:2604.07595v1 Announce Type: cross Abstract: Language model agents reason from scratch on every query: each time an agent retrieves evidence and deliberates, the chain of thought is discarded and

Restoring Heterogeneity in LLM-based Social Simulation: An Audience Segmentation Approach

Model ReleasesDGX agent

arXiv:2604.06663v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to simulate social attitudes and behaviors, offering scalable 'silicon samples' that can approximat

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models

Model ReleasesDGX agent

arXiv:2604.06996v1 Announce Type: cross Abstract: LLM-as-a-judge has become the de facto approach for evaluating LLM outputs. However, judges are known to exhibit self-preference bias (SPB): they tend

SentinelSphere: Integrating AI-Powered Real-Time Threat Detection with Cybersecurity Awareness Training

Model ReleasesDGX agent

arXiv:2604.06900v1 Announce Type: cross Abstract: The field of cybersecurity is confronted with two interrelated challenges: a worldwide deficit of qualified practitioners and ongoing human-factor wea

SERHANT.'s playbook for rapid AI iteration

ToolsDGX agent

SERHANT., an AI-native real estate brokerage, built its internal AI platform S.MPLE on Next.js and Vercel, scaling it from a 200-agent pilot to over 900 agents without replatforming or rebuilding i...

SIM1: Physics-Aligned Simulator as Zero-Shot Data Scaler in Deformable Worlds

SafetyDGX agent

arXiv:2604.08544v1 Announce Type: cross Abstract: Robotic manipulation with deformable objects represents a data-intensive regime in embodied learning, where shape, contact, and topology co-evolve in

SkillSieve: A Hierarchical Triage Framework for Detecting Malicious AI Agent Skills

Model ReleasesDGX agent

arXiv:2604.06550v1 Announce Type: cross Abstract: OpenClaw's ClawHub marketplace hosts over 13,000 community-contributed agent skills, and between 13% and 26% of them contain security vulnerabilities

Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation

Model ReleasesDGX agent

arXiv:2604.07028v1 Announce Type: cross Abstract: Strategic interaction in adversarial domains such as law, diplomacy, and negotiation is mediated by language, yet most game-theoretic models abstract

SubSearch: Intermediate Rewards for Unsupervised Guided Reasoning in Complex Retrieval

AgentsDGX agent

arXiv:2604.07415v1 Announce Type: cross Abstract: Large language models (LLMs) are probabilistic in nature and perform more reliably when augmented with external information. As complex queries often

@swyx 👀 new Chief AI Officer · UK Government 👀

ToolsDGX agent

The UK government hired Kalbir Sohi to the newly created role of Chief AI Officer (CAIO), making it the most senior AI leadership position in the country's public sector. Sohi holds the title of ...

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

Model ReleasesDGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

The Impact of Steering Large Language Models with Persona Vectors in Educational Applications

Model ReleasesDGX agent

arXiv:2604.07102v1 Announce Type: cross Abstract: Activation-based steering can personalize large language models at inference time, but its effects in educational settings remain unclear. We study pe

tldr > evals are the new training data. instead of updating weights, you're updating the agent harness > problem is agents are famous cheate…

AgentsDGX agent

tldr > evals are the new training data. instead of updating weights, you're updating the agent harness > problem is agents are famous cheaters. they will reward-hack your evals and overfit just to mak

today, a tall guy in a colorful sweater walked up to me. i was already at the end of my social energy reserves, having dozens of people walk…

ToolsDGX agent

today, a tall guy in a colorful sweater walked up to me. i was already at the end of my social energy reserves, having dozens of people walk up to me, never having 5 minutes to breath. the guy just wa

TREASURE: The Visa Payment Foundation Model for High-Volume Transaction Understanding

ApplicationsDGX agent

arXiv:2511.19693v3 Announce Type: replace-cross Abstract: Payment networks form the backbone of modern commerce, generating high volumes of transaction records from daily activities. Properly modeling

VertAX: a differentiable vertex model for learning epithelial tissue mechanics

Model ReleasesDGX agent

arXiv:2604.06896v1 Announce Type: new Abstract: Epithelial tissues dynamically reshape through local mechanical interactions among cells, a process well captured by vertex models. Yet their many tunab

Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment

ApplicationsDGX agent

arXiv:2604.08212v1 Announce Type: new Abstract: General-purpose vision-language models demonstrate strong performance in everyday domains but struggle with specialized technical fields requiring preci

When to Call an Apple Red: Humans Follow Introspective Rules, VLMs Don't

Model ReleasesDGX agent

arXiv:2604.06422v1 Announce Type: cross Abstract: Understanding when Vision-Language Models (VLMs) will behave unexpectedly, whether models can reliably predict their own behavior, and if models adher

9 Apr 2026

@3blue1brown https://x.com/verncrawford/status/2042344703077859402?s=46

AgentsDGX agent

@3blue1brown https://x.com/verncrawford/status/2042344703077859402?s=46 @NousResearch 2 prompts guys. 2 prompts!!! The second prompt was “go from 6 camera moves to 3.” (Hermes Agent, Codex5.4, Manim V

a useful mental model on how teams can think about good data design to improve their models/agents: Evals ~= Training Data ~= Environments -…

AgentsDGX agent

a useful mental model on how teams can think about good data design to improve their models/agents: Evals ~= Training Data ~= Environments - in Classical Deep Learning, we learn from each training exa

Announcing the @LangChain podcast -- Max Agency. Deep context to give you the edge while building and iterating on agents. Watch the full ep…

ApplicationsDGX agent

Announcing the @LangChain podcast -- Max Agency. Deep context to give you the edge while building and iterating on agents. Watch the full episode on: - Youtube: https://www.youtube.com/watch?v=Xyh1Eqc

b8737

Local AiDGX agent

llama.cpp release **b8737** is a focused maintenance build that adds missing CUDA error handling to the ggml backend. Specifically, it checks the return values of NVIDIA CUB library calls used in t...

calling all of my fellow podcast nerds!

ApplicationsDGX agent

calling all of my fellow podcast nerds! 🎙️Introducing Max Agency Max Agency is a new podcast where we go deep on how the best agents are actually being built: architecture decisions, tradeoffs, evals,

Can’t get enough of Harrison and his AI takes? Lucky for you, we’re announcing Max Agency where you can hear him & other industry experts ta…

ApplicationsDGX agent

Can’t get enough of Harrison and his AI takes? Lucky for you, we’re announcing Max Agency where you can hear him & other industry experts talk agents, harnesses, models, evals, and more!! 🎙️Introducin

Checkout our new Max Agency podcast - the first episode is already out!

ApplicationsDGX agent

Checkout our new Max Agency podcast - the first episode is already out! 🎙️Introducing Max Agency Max Agency is a new podcast where we go deep on how the best agents are actually being built: architect

Day 1 of @aiDotEngineer in London – what a ride! 🤖 Ash Prabaker & Andrew Wilson from Anthropic on 'How to Build Agents That Run for Hours (…

Model ReleasesDGX agent

Day 1 of @aiDotEngineer in London – what a ride! 🤖 Ash Prabaker & Andrew Wilson from Anthropic on 'How to Build Agents That Run for Hours (Without Losing the Plot)' – spoiler: stop letting your agents

Elon talks about how Starship was built, and it really puts things into perspective This is 'probably the biggest thing ever made by pure hu…

IndustryDGX agent

Elon talks about how Starship was built, and it really puts things into perspective This is 'probably the biggest thing ever made by pure human hands' • Created entirely pre-AI • 'At the limit of biol

extremely fun jamming with Harrison on all things analytics, agents and evals! I have like 20 more hours worth of takes and opinions in this…

Model ReleasesDGX agent

extremely fun jamming with Harrison on all things analytics, agents and evals! I have like 20 more hours worth of takes and opinions in this space and going on the pod uncorked them, more to come for

Great morning bringing the speakers from @aiDotEngineer to Downing Street to discuss transforming the state. Through the Incubator for AI an…

ToolsDGX agent

Great morning bringing the speakers from @aiDotEngineer to Downing Street to discuss transforming the state. Through the Incubator for AI and the No10 Innovation Fellowship, we are making sure that to

← Previous
1…87888990
Next →