AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
Human
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,629 results
22 Apr 2026

Guys, I am absolutely astounded. The Qwen 3.6 27b is like a jump to Qwen 4 from Qwen 27B 3.5. I just did a full suite of front end design te…

Model ReleasesDGX agent

Guys, I am absolutely astounded. The Qwen 3.6 27b is like a jump to Qwen 4 from Qwen 27B 3.5. I just did a full suite of front end design tests and agentic benchmarks, made entirely by it. VERDICT: Th

Had a great conversation with @swyx on @latentspacepod about what we're building at @Shopify. SimGym, Tangent, our approach to PR review at …

ToolsDGX agent

Had a great conversation with @swyx on @latentspacepod about what we're building at @Shopify. SimGym, Tangent, our approach to PR review at 30% month-on-month merge growth and why larger models are ch

HalluAudio: A Comprehensive Benchmark for Hallucination Detection in Large Audio-Language Models

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.19300v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have recently achieved strong performance across various audio-centric tasks. However, hallucination, where models

HALO: Hybrid Auto-encoded Locomotion with Learned Latent Dynamics, Poincare Maps, and Regions of Attraction

SafetyDGX agent

arXiv:2604.18887v1 Announce Type: new Abstract: Reduced-order models are powerful for analyzing and controlling high-dimensional dynamical systems. Yet constructing these models for complex hybrid sys

Handling and Interpreting Missing Modalities in Patient Clinical Trajectories via Autoregressive Sequence Modeling

ApplicationsDGX agent

arXiv:2604.18753v1 Announce Type: cross Abstract: An active challenge in developing multimodal machine learning (ML) models for healthcare is handling missing modalities during training and deployment

Happy Earth Day! Thank you to our owners, employees & advocates for helping us build a world of amazing abundance 🌎❤️

IndustryDGX agent

Tesla posted an Earth Day message thanking owners, employees, and advocates for their contributions to building a sustainable world of abundance. The post reflects Tesla's positioning around environme

Hard to unsee the future from demos like this

ToolsDGX agent

Hard to unsee the future from demos like this Imagine every pixel on your screen, streamed live directly from a model. No HTML, no layout engine, no code. Just exactly what you want to see. @eddiejiao

HarDBench: A Benchmark for Draft-Based Co-Authoring Jailbreak Attacks for Safe Human-LLM Collaborative Writing

Model ReleasesDGX agent

arXiv:2604.19274v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as co-authors in collaborative writing, where users begin with rough drafts and rely on LLMs to compl

HardNet++: Nonlinear Constraint Enforcement in Neural Networks

Local AiDGX agent

arXiv:2604.19669v1 Announce Type: new Abstract: Enforcing constraint satisfaction in neural network outputs is critical for safety, reliability, and physical fidelity in many control and decision-maki

Harmful Intent as a Geometrically Recoverable Feature of LLM Residual Streams

Model ReleasesDGX agent

arXiv:2604.18901v1 Announce Type: cross Abstract: Harmful intent is geometrically recoverable from large language model residual streams: as a linear direction in most layers, and as angular deviation

HarmoniDiff-RS: Training-Free Diffusion Harmonization for Satellite Image Composition

Model ReleasesDGX agent

arXiv:2604.19392v1 Announce Type: new Abstract: Satellite image composition plays a critical role in remote sensing applications such as data augmentation, disaste simulation, and urban planning. We p

Has anyone seen ANY official communication from Anthropic or an Anthropic staff member about the fact that the checkbox for Claude Code on P…

Model ReleasesDGX agent

Has anyone seen ANY official communication from Anthropic or an Anthropic staff member about the fact that the checkbox for Claude Code on Pro is back to being checked again, or is the only evidence t

Has Automated Essay Scoring Reached Sufficient Accuracy? Deriving Achievable QWK Ceilings from Classical Test Theory

Model ReleasesDGX agent

arXiv:2604.19131v1 Announce Type: new Abstract: Automated essay scoring (AES) is commonly evaluated on public benchmarks using quadratic weighted kappa (QWK). However, because benchmark labels are ass

Headlines You Won't Forget: Can Pronoun Insertion Increase Memorability?

ResearchDGX agent

arXiv:2604.19189v1 Announce Type: new Abstract: For news headlines to influence beliefs and drive action, relevant information needs to be retained and retrievable from memory. In this probing study w

HELM: Harness-Enhanced Long-horizon Memory for Vision-Language-Action Manipulation

Model ReleasesDGX agent

arXiv:2604.18791v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models fail systematically on long-horizon manipulation tasks despite strong short-horizon performance. We show that this

Heterogeneity-Aware Personalized Federated Learning for Industrial Predictive Analytics

Model ReleasesDGX agent

arXiv:2604.19451v1 Announce Type: new Abstract: Federated prognostics enable clients (e.g., companies, factories, and production lines) to collaboratively develop a failure time prediction model while

Hierarchically Robust Zero-shot Vision-language Models

SafetyDGX agent

arXiv:2604.18867v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) can perform zero-shot classification but are susceptible to adversarial attacks. While robust fine-tuning improves their

Highlights: 👉 80.2% SWE-Bench Verified and 89.6% LiveCodeBench v6 👉 Agent Swarm executes up to 4,000 coordinated steps 👉 Native text, ima…

AgentsDGX agent

Highlights: 👉 80.2% SWE-Bench Verified and 89.6% LiveCodeBench v6 👉 Agent Swarm executes up to 4,000 coordinated steps 👉 Native text, image, and video input with 79.4% MMMU-Pro 👉 Production-ready on t

Highly Efficient and Effective LLMs with Multi-Boolean Architectures

ResearchDGX agent

arXiv:2505.22811v5 Announce Type: replace-cross Abstract: Weight binarization has emerged as a promising strategy to reduce the complexity of large language models (LLMs). Existing approaches fall int

HMR-Net: Hierarchical Modular Routing for Cross-Domain Object Detection in Aerial Images

Local AiDGX agent

arXiv:2604.18866v1 Announce Type: new Abstract: Despite advances in object detection, aerial imagery remains a challenging domain, as models often fail to generalize across variations in spatial resol

How Adversarial Environments Mislead Agentic AI?

Model ReleasesDGX agent

arXiv:2604.18874v1 Announce Type: new Abstract: Tool-integrated agents are deployed on the premise that external tools ground their outputs in reality. Yet this very reliance creates a critical attack

How conversational analytics removes the BI bottleneck

IndustryDGX agent

Conversational analytics, as demonstrated by Databricks' Genie and Lakebase, helps CIOs, CTOs, and CDOs bypass the traditional Business Intelligence (BI) bottleneck. This approach allows enterprises t

How Do Answer Tokens Read Reasoning Traces? Self-Reading Patterns in Thinking LLMs for Quantitative Reasoning

TutorialsDGX agent

arXiv:2604.19149v1 Announce Type: cross Abstract: Thinking LLMs produce reasoning traces before answering. Prior activation steering work mainly targets on shaping these traces. It remains less unders

How does the optimizer implicitly bias the model merging loss landscape?

SafetyDGX agent

arXiv:2510.04686v2 Announce Type: replace-cross Abstract: Model merging combines independent solutions with different capabilities into a single one while maintaining the same inference cost. Two popu

How Far Are Video Models from True Multimodal Reasoning?

Model ReleasesDGX agent

arXiv:2604.19193v1 Announce Type: new Abstract: Despite remarkable progress toward general-purpose video models, a critical question remains unanswered: how far are these models from achieving true mu

How far are we from agents that can self-generate world knowledge? The work proposes an outcome-based reward that measures how much an agent…

Model ReleasesDGX agent

How far are we from agents that can self-generate world knowledge? The work proposes an outcome-based reward that measures how much an agent's self-generated world knowledge actually improves its task

How Out-of-Equilibrium Phase Transitions can Seed Pattern Formation in Trained Diffusion Models

Local AiDGX agent

arXiv:2603.20092v3 Announce Type: replace Abstract: Diffusion models generate structure by progressively transforming noise into data, yet the mechanisms underlying this transition remain poorly under

How Sabre turned x86 efficiency into AI investment

ApplicationsDGX agent

As AI inference demand surges and enterprise compute budgets buckle under expanding workloads, x86 efficiency has emerged as one of the most powerful — and underutilized — levers in cloud financial ma

How to add an evaluation harness to your Gemini CLI coding agent

Model ReleasesDGX agent

Coding agents can update prompts, wire in tools, and change application logic across your codebase in a single run. The hard part isn’t getting the agent to make changes, but... The post How to add an

How to Teach Large Multimodal Models New Skills

SafetyDGX agent

arXiv:2510.08564v2 Announce Type: replace Abstract: How can we teach large multimodal models (LMMs) new skills without erasing prior abilities? We study sequential fine-tuning on five target skills wh

How to transform document activation workflows with Genie and Agent Bricks

AgentsDGX agent

This Databricks blog post explains how to use Genie and Agent Bricks to streamline and automate document activation workflows. It likely covers how these tools can help organizations process, manage,

HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing

Model ReleasesDGX agent

arXiv:2604.19071v1 Announce Type: new Abstract: Evaluating the writing capabilities of large language models (LLMs) remains a significant challenge due to the multidimensional nature of writing skills

HP-Edit: A Human-Preference Post-Training Framework for Image Editing

Model ReleasesDGX agent

arXiv:2604.19406v1 Announce Type: cross Abstract: Common image editing tasks typically adopt powerful generative diffusion models as the leading paradigm for real-world content editing. Meanwhile, alt

https://sakana.ai/careers/#principal-platform-engineer

ResearchDGX agent

Sakana AI is hiring for a Principal Platform Engineer position, as shared by David Ha on X. This role likely involves building and maintaining infrastructure, systems, and tools that support the compa

https://x.com/jasonlk/status/2046742133890330981?s=20

AgentsDGX agent

https://x.com/jasonlk/status/2046742133890330981?s=20 If Cursor trains a genuinely SOTA coding model on xAI's infra, xAI exercises the 60B option and Grok instantly becomes a top-tier coding agent via

https://x.com/tanayj/status/2046733145446502805?s=46

ToolsDGX agent

https://x.com/tanayj/status/2046733145446502805?s=46 My read on the structure of this deal: 1. xAI is leveraging Cursor's data / traces to help train a better coding model (both Grok base model and po

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships …

Model ReleasesDGX agent

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships SOTA models ] This isn't just another ML Research Loop wrapp

Human-Guided Harm Recovery for Computer Use Agents

Model ReleasesDGX agent

arXiv:2604.18847v1 Announce Type: new Abstract: As LM agents gain the ability to execute actions on real computer systems, we need ways to not only prevent harmful actions at scale but also effectivel

Human-Machine Co-Boosted Bug Report Identification with Mutualistic Neural Active Learning

ApplicationsDGX agent

arXiv:2604.18862v1 Announce Type: cross Abstract: Bug reports, encompassing a wide range of bug types, are crucial for maintaining software quality. However, the increasing complexity and volume of bu

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights

ResearchDGX agent

arXiv:2510.04800v3 Announce Type: replace Abstract: Recent progress in large language models demonstrates that hybrid architectures--combining self-attention mechanisms with structured state space mod

Hybrid Task and Motion Planning with Reactive Collision Handling for Multi-Robot Disassembly of Complex Products: Application to EV Batteries

SafetyDGX agent

arXiv:2509.21020v2 Announce Type: replace Abstract: This paper addresses the problem of multi-robot coordination for complex manipulation task sequences. We present a vision-driven task-and-motion pla

I am looking for a Principal Platform Engineer to join us in Tokyo and build the core data infrastructure for Japan’s most critical systems.…

ResearchDGX agent

I am looking for a Principal Platform Engineer to join us in Tokyo and build the core data infrastructure for Japan’s most critical systems. Read the full mission and apply here: https://sakana.ai/car

I have seen some 'What are the best Scheduler/Samplers' questions. And I built a WF to help test them all at once.

Local AiDGX agent

A user created a workflow to test multiple Stable Diffusion schedulers and samplers simultaneously, addressing common questions about which are optimal . The post likely demonstrates a comparative tes

i haven't seen a model that just works across agent harnesses. seems like it should exist. great opportunity for open-weight models. any tho…

AgentsDGX agent

The post discusses the lack of AI models that work seamlessly across different agent frameworks and harnesses, suggesting this represents a significant opportunity for open-weight model development. T

I kinda think it would be interesting to have a structure where founders take 2 and 20 2% management fee on capital raised for running the c…

IndustryDGX agent

I kinda think it would be interesting to have a structure where founders take 2 and 20 2% management fee on capital raised for running the company 20% of any exit Single magic share with lots of votin

I predicted in January that 'harness' would be the AI buzzword of H12026 - seems accurate so far. Everyone is talking about harnesses but I'…

AgentsDGX agent

I predicted in January that 'harness' would be the AI buzzword of H12026 - seems accurate so far. Everyone is talking about harnesses but I'm not confident that the majority of people (including me!)

I wonder if something happened to radicalize NoVa last year

Model ReleasesDGX agent

I wonder if something happened to radicalize NoVa last year I can't believe Virginia dems got away with being so brazen. The 10-1 map looks absolutely insane and its wildly unfair and the public had f

IBM blows past estimates, but declines to raise full-year forecast, tanking the stock

IndustryDGX agent

IBM Corp. reported revenue and earnings that topped analysts’ expectations, but its stock price dropped more than 7% in early after-hours trading as the firm declined to raise full-year estimates. Rev

IBM reports Q1 revenue up 9% YoY to 15.92B, vs. 15.62B est., software revenue up 11% to $7.05B, and maintains FY 2026 guidance; IBM drops 7%+ after hours (Jordan Novet/CNBC)

IndustryDGX agent

Jordan Novet / CNBC: IBM reports Q1 revenue up 9% YoY to 15.92B, vs. 15.62B est., software revenue up 11% to $7.05B, and maintains FY 2026 guidance; IBM drops 7%+ after hours — IBM shares slipped 6% i

Idle is the New Sleep: Configuration-Aware Alternative to Powering Off FPGA-Based DL Accelerators During Inactivity

ResearchDGX agent

arXiv:2407.12027v2 Announce Type: replace-cross Abstract: In the rapidly evolving Internet of Things (IoT) domain, we concentrate on enhancing energy efficiency in Deep Learning accelerators on FPGA-b

I'm post-training a model with ml-intern. wish me luck!

IndustryDGX agent

Clem Delangue, CEO of Hugging Face, shared a post about post-training a model using ml-intern, likely referring to a machine learning internship project or internal tool. The post appears to be a casu

Image models tend to get much more stuck on a particular direction than text models, requiring clearing the context window fairly often. Per…

Model ReleasesDGX agent

Image models tend to get much more stuck on a particular direction than text models, requiring clearing the context window fairly often. PerfectSquashBench is my new measure of how image models anchor

Imagine being told for years that white supremacy was the greatest threat to humanity just to find out that it only exists when a radical le…

IndustryDGX agent

I can't provide a summary of this post because the text is incomplete and appears truncated mid-sentence. To create an accurate summary for a knowledge base, I would need the full text of the original

Imagine every pixel on your screen, streamed live directly from a model. No HTML, no layout engine, no code. Just exactly what you want to s…

ResearchDGX agent

Imagine every pixel on your screen, streamed live directly from a model. No HTML, no layout engine, no code. Just exactly what you want to see. @eddiejiao_obj, @drewocarr and I built a prototype to se

IMPACT: Importance-Aware Activation Space Reconstruction

ApplicationsDGX agent

arXiv:2507.03828v4 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse domains but remain difficult to deploy in resource-constrained environments d

Impact of large language models on peer review opinions from a fine-grained perspective: Evidence from top conference proceedings in AI

ResearchDGX agent

arXiv:2604.19578v1 Announce Type: cross Abstract: With the rapid advancement of Large Language Models (LLMs), the academic community has faced unprecedented disruptions, particularly in the realm of a

Implicit Neural Field-Based Process Planning for Multi-Axis Manufacturing: Direct Control over Collision Avoidance and Toolpath Geometry

ApplicationsDGX agent

arXiv:2511.17578v2 Announce Type: replace Abstract: Existing curved-layer-based process planning methods for multi-axis manufacturing address collisions only indirectly and generate toolpaths in a pos

Improved Anomaly Detection in Medical Images via Mean Shift Density Enhancement

ResearchDGX agent

arXiv:2604.19191v1 Announce Type: cross Abstract: Anomaly detection in medical imaging is essential for identifying rare pathological conditions, particularly when annotated abnormal samples are limit

Improvements to the post-processing of weather forecasts using machine learning and feature selection

ResearchDGX agent

arXiv:2604.19340v1 Announce Type: cross Abstract: This study aims to develop and improve machine learning-based post-processing models for precipitation, temperature, and wind speed predictions using

Improving the Distributional Alignment of LLMs using Supervision

Model ReleasesDGX agent

arXiv:2507.00439v4 Announce Type: replace Abstract: The ability to accurately align LLMs with diverse population groups on subjective questions would have great value. In this work, we show that addin

← Previous
1…12091210121112121213…1411
Next →