AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,832 results
14 Apr 2026

WebLLM: A High-Performance In-Browser LLM Inference Engine

Local AiDGX agent

arXiv:2412.15803v2 Announce Type: replace-cross Abstract: Advancements in large language models (LLMs) have unlocked remarkable capabilities. While deploying these models typically requires server-gra

Weird Generalization is Weirdly Brittle

SafetyDGX agent

arXiv:2604.10022v1 Announce Type: new Abstract: Weird generalization is a phenomenon in which models fine-tuned on data from a narrow domain (e.g. insecure code) develop surprising traits that manifes

We’re expanding Trusted Access for Cyber with additional tiers for authenticated cybersecurity defenders. Customers in the highest tiers can…

Model ReleasesDGX agent

We’re expanding Trusted Access for Cyber with additional tiers for authenticated cybersecurity defenders. Customers in the highest tiers can request access to GPT-5.4-Cyber, a version of GPT-5.4 fine-

Content type
AllBlogX PostPaperYouTubeRedditGitHub

We're Transferring the Stripe Sync Engine to Stripe

ToolsDGX agent

Supabase originally built the open-source Stripe Sync Engine to keep its own Postgres database of billing data current with Stripe, using webhooks to update records for customers, invoices, and paymen

We've been building something that doesn't fit in a wave. Coming soon

ToolsDGX agent

Windsurf, the AI-powered coding platform, has teased an upcoming product or feature announcement that is described as something that doesn't fit within their existing 'wave' release framework. The cry

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to ap…

HardwareDGX agent

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to apply it to optimizing CUDA kernels. In 3 weeks, it delivered

We've been working on this for a while. Can't wait to hear what you think

Model ReleasesDGX agent

We've been working on this for a while. Can't wait to hear what you think We've redesigned Claude Code on desktop. You can now run multiple Claude sessions side by side from one window, with a new sid

We’ve entered the era of AI

IndustryDGX agent

We’ve entered the era of AI MPA boss Charles Rivkin says AI can “bolster the art of storytelling” and “improve the fan experience”: “We’ve entered the era of AI,” Rivkin told theater operators at #Cin

We've tested new OSS models the moment they're released for a while at Lindy. Inference is our #1 cost by a lot (more than payroll) — cuttin…

Model ReleasesDGX agent

We've tested new OSS models the moment they're released for a while at Lindy. Inference is our #1 cost by a lot (more than payroll) — cutting it by 2-5x would be transformative. Last year, OSS models

What and Where to Adapt: Structure-Semantics Co-Tuning for Machine Vision Compression via Synergistic Adapters

Model ReleasesDGX agent

arXiv:2604.10017v1 Announce Type: new Abstract: Parameter-efficient fine-tuning of pre-trained codecs is a promising direction in image compression for human and machine vision. While most existing wo

What are some good env versions for speed and compatibility?

Local AiDGX agent

This Reddit thread from r/StableDiffusion discusses community recommendations for optimal Python, CUDA, PyTorch, and related dependency versions when setting up a Stable Diffusion environment that bal

What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?

ApplicationsDGX agent

arXiv:2604.11374v1 Announce Type: cross Abstract: Personalized image aesthetics assessment (PIAA) is an important research problem with practical real-world applications. While methods based on vision

What do your logits know? (The answer may surprise you!)

TutorialsDGX agent

arXiv:2604.09885v1 Announce Type: new Abstract: Recent work has shown that probing model internals can reveal a wealth of information not apparent from the model generations. This poses the risk of un

What Factors Affect LLMs and RLLMs in Financial Question Answering?

SafetyDGX agent

arXiv:2507.08339v4 Announce Type: replace Abstract: Recently, large language models (LLMs) and reasoning large language models (RLLMs) have gained considerable attention from many researchers. RLLMs e

What is the AC guidance for ICML? (Or: ICML qq thread) [D]

ResearchDGX agent

This r/MachineLearning Reddit thread serves as a community Q&A ('qq thread') focused on the Area Chair (AC) process for ICML, where researchers ask and answer questions about AC guidance, responsibili

What I’ve been building: ATOM Report, post-training course, finishing my book, and ongoing research

TutorialsDGX agent

The author provided an update on several ongoing technical projects, including the ATOM Report, a new post-training course, and

What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction

Model ReleasesDGX agent

arXiv:2407.08101v4 Announce Type: replace Abstract: Vision-language models have shown impressive progress in recent years. However, existing models are largely limited to turn-based interactions, wher

What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.06165v2 Announce Type: replace-cross Abstract: Current vision-language benchmarks predominantly feature well-structured questions with clear, explicit prompts. However, real user queries ar

What's going on?

IndustryDGX agent

This Reddit post from r/ChatGPT, titled 'What's going on?', likely reflects user reactions and discussion around a notable change or issue with ChatGPT — such as the removal of specific models, unexpe

What's In My Human Feedback? Learning Interpretable Descriptions of Preference Data

SafetyDGX agent

arXiv:2510.26202v2 Announce Type: replace-cross Abstract: Human feedback can alter language models in unpredictable and undesirable ways, as practitioners lack a clear understanding of what feedback d

what's the best place to buy GPU server?

HardwareDGX agent

This Reddit thread from r/ollama discusses community recommendations for purchasing or renting GPU servers to run Ollama and local LLMs. It likely covers options ranging from dedicated GPU server prov

When Can You Poison Rewards? A Tight Characterization of Reward Poisoning in Linear MDPs

SafetyDGX agent

arXiv:2604.10062v1 Announce Type: new Abstract: We study reward poisoning attacks in reinforcement learning (RL), where an adversary manipulates rewards within constrained budgets to force the target

When did ChatGPT get a clock?

IndustryDGX agent

A Reddit thread on r/ChatGPT where users noticed and discussed ChatGPT apparently gaining the ability to report the current time — a capability it historically lacked. The base ChatGPT model has no bu

when duplicating the chatgpt logo it forms a symmetric star pattern is this just geometri

IndustryDGX agent

A viral observation circulating on Reddit and social media notes that when the ChatGPT logo is duplicated and mirrored, it appears to form a symmetric star pattern — specifically a dodecagram with a s

When Meaning Isn't Literal: Exploring Idiomatic Meaning Across Languages and Modalities

Model ReleasesDGX agent

arXiv:2604.10787v1 Announce Type: new Abstract: Idiomatic reasoning, deeply intertwined with metaphor and culture, remains a blind spot for contemporary language models, whose progress skews toward su

When More Thinking Hurts: Overthinking in LLM Test-Time Compute Scaling

ResearchDGX agent

arXiv:2604.10739v1 Announce Type: new Abstract: Scaling test-time compute through extended chains of thought has become a dominant paradigm for improving large language model reasoning. However, exist

When simulations look right but causal effects go wrong: Large language models as behavioral simulators

SafetyDGX agent

arXiv:2604.02458v2 Announce Type: replace-cross Abstract: Behavioral simulation is increasingly used to anticipate responses to interventions. Large language models (LLMs) enable researchers to specif

When Valid Signals Fail: Regime Boundaries Between LLM Features and RL Trading Policies

SafetyDGX agent

arXiv:2604.10996v1 Announce Type: cross Abstract: Can large language models (LLMs) generate continuous numerical features that improve reinforcement learning (RL) trading agents? We build a modular pi

When Verification Fails: How Compositionally Infeasible Claims Escape Rejection

ResearchDGX agent

arXiv:2604.10990v1 Announce Type: cross Abstract: Scientific claim verification, the task of determining whether claims are entailed by scientific evidence, is fundamental to establishing discoveries

Who Gets Which Message? Auditing Demographic Bias in LLM-Generated Targeted Text

Model ReleasesDGX agent

arXiv:2601.17172v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly capable of generating personalized, persuasive text at scale, raising new questions about bias a

Who Handles Orientation? Investigating Invariance in Feature Matching

TutorialsDGX agent

arXiv:2604.11809v1 Announce Type: new Abstract: Finding matching keypoints between images is a core problem in 3D computer vision. However, modern matchers struggle with large in-plane rotations. A st

Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry

Model ReleasesDGX agent

arXiv:2604.10101v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AI-generated literary creati

Who's coming to AI Engineer Miami? @swyx and @liamcbride will be there! Grab the last few available tickets and meet us there!

ToolsDGX agent

Who's coming to AI Engineer Miami? @swyx and @liamcbride will be there! Grab the last few available tickets and meet us there! We are 10 days away! Join us in Miami and see incredible speakers like: @

Who's going to be at AIE Miami! I wanna hang with you.

ToolsDGX agent

This appears to be a social media post from Swyx (likely Shawn Wang, a developer advocate and AI enthusiast) on X (formerly Twitter), seeking to connect with other attendees at an AI Engineer (AIE) ev

Why Code, Why Now: An Information-Theoretic Perspective on the Limits of Machine Learning

Local AiDGX agent

arXiv:2602.13934v4 Announce Type: replace-cross Abstract: This paper offers a new perspective on the limits of machine learning: the ceiling on progress is set not by model size or algorithm choice bu

Why do GPT responses feel exhausting even when they’re right?

IndustryDGX agent

A Reddit thread from r/ChatGPT exploring the user experience phenomenon where GPT responses can feel cognitively draining or over-engineered even when technically accurate — touching on issues like ve

Why Do Large Language Models Generate Harmful Content?

ResearchDGX agent

arXiv:2604.11663v1 Announce Type: new Abstract: Large Language Models (LLMs) have been shown to generate harmful content. However, the underlying causes of such behavior remain under explored. We prop

Why Do Multilingual Reasoning Gaps Emerge in Reasoning Language Models?

ResearchDGX agent

arXiv:2510.27269v3 Announce Type: replace-cross Abstract: Reasoning language models (RLMs) achieve strong performance on complex reasoning tasks, yet they still exhibit a multilingual reasoning gap, p

Why does the upgrade button looks like Gemini logo 🤨

Model ReleasesDGX agent

This Reddit post from r/ChatGPT is a user-generated discussion in which a member notices that ChatGPT's upgrade button visually resembles Google's Gemini logo — a star-like, multi-pointed sparkle icon

Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs

ApplicationsDGX agent

arXiv:2604.10495v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in real-world applications, reliable uncertainty quantification (UQ) becomes critical for safe

Why I just quit Claude Pro after 48 hours (Rate Limit Anxiety)

Model ReleasesDGX agent

A Reddit post on r/ChatGPT in which a user describes canceling their Claude Pro subscription within 48 hours, citing frustration with Claude's usage rate limits as the primary reason. The post reflect

Why Smaller Is Slower? Dimensional Misalignment in Compressed LLMs

Model ReleasesDGX agent

arXiv:2604.09595v1 Announce Type: cross Abstract: Post-training compression reduces LLM parameter counts but often produces irregular tensor dimensions that degrade GPU performance -- a phenomenon we

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

Model ReleasesDGX agent

arXiv:2602.02343v3 Announce Type: replace-cross Abstract: Methods for controlling large language models (LLMs), including local weight fine-tuning, LoRA-based adaptation, and activation-based interven

Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.10079v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) is the standard approach for adapting large language models (LLMs) to downstream tasks. However, we observe a persistent fa

Why why why is it so hard to understand that stupid systems can still be dangerous?

SafetyDGX agent

Why why why is it so hard to understand that stupid systems can still be dangerous? @Graffitinights1 For the zillionth time, dumb systems that are empowered can dangerous. like this monstrosity, as an

WiFlow: A Lightweight WiFi-based Continuous Human Pose Estimation Network with Spatio-Temporal Feature Decoupling

ApplicationsDGX agent

arXiv:2602.08661v2 Announce Type: replace Abstract: Human pose estimation is fundamental to intelligent perception in the Internet of Things (IoT), enabling applications ranging from smart healthcare

WisPaper: Your AI Scholar Search Engine

AgentsDGX agent

arXiv:2512.06879v3 Announce Type: replace-cross Abstract: We present extsc{WisPaper}, an end-to-end agent system that transforms how researchers discover, organize, and track academic literature. The

WM-DAgger: Enabling Efficient Data Aggregation for Imitation Learning with World Models

SafetyDGX agent

arXiv:2604.11351v1 Announce Type: new Abstract: Imitation learning is a powerful paradigm for training robotic policies, yet its performance is limited by compounding errors: minor policy inaccuracies

Wolkowicz-Styan Upper Bound on the Hessian Eigenspectrum for Cross-Entropy Loss in Nonlinear Smooth Neural Networks

ResearchDGX agent

arXiv:2604.10202v1 Announce Type: cross Abstract: Neural networks (NNs) are central to modern machine learning and achieve state-of-the-art results in many applications. However, the relationship betw

WOODELF-HD: Efficient Background SHAP for High-Depth Decision Trees

ResearchDGX agent

arXiv:2604.10569v1 Announce Type: new Abstract: Decision-tree ensembles are a cornerstone of predictive modeling, and SHAP is a standard framework for interpreting their predictions. Among its variant

Woosh: A Sound Effects Foundation Model

Model ReleasesDGX agent

arXiv:2604.01929v2 Announce Type: replace-cross Abstract: The audio research community depends on open generative models as foundational tools for building novel approaches and establishing baselines.

Wordle 1,759 3/6 ⬛⬛⬛⬛🟨 🟨🟨🟨⬛⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

Anthropic's official X (Twitter) account shared a Wordle result showing the puzzle was solved in 3 out of 6 attempts for Wordle #1,759. The grid reveals the solver had no correct letters in the first

Work in Progress Encoder and Decoder!

Local AiDGX agent

A Reddit post from the r/StableDiffusion community sharing a work-in-progress development of a custom encoder and decoder, likely relating to the VAE (Variational Autoencoder) or latent space componen

Working Paper: Towards Schema-based Learning from a Category-Theoretic Perspective

AgentsDGX agent

arXiv:2604.10589v1 Announce Type: new Abstract: We introduce a hierarchical categorical framework for Schema-Based Learning (SBL) structured across four interconnected levels. At the schema level, a f

X-SYS: A Reference Architecture for Interactive Explanation Systems

ResearchDGX agent

arXiv:2602.12748v3 Announce Type: replace Abstract: The explainable AI (XAI) research community has proposed numerous technical methods, yet deploying explainability as systems remains challenging: In

XD-MAP: Cross-Modal Domain Adaptation via Semantic Parametric Maps for Scalable Training Data Generation

ResearchDGX agent

arXiv:2601.14477v2 Announce Type: replace-cross Abstract: Until open-world foundation models match the performance of specialized approaches, deep learning systems remain dependent on task- and sensor

YIELD: A Large-Scale Dataset and Evaluation Framework for Information Elicitation Agents

SafetyDGX agent

arXiv:2604.10968v1 Announce Type: new Abstract: Most conversational agents (CAs) are designed to satisfy user needs through user-driven interactions. However, many real-world settings, such as academi

You can decompose models into a graph database [N]

ResearchDGX agent

This Reddit post from r/MachineLearning discusses the concept of decomposing machine learning models into a graph database representation, treating a model's components — such as layers, weights, and

You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass

Model ReleasesDGX agent

arXiv:2604.10966v1 Announce Type: cross Abstract: We present a discriminative multimodal reward model that scores all candidate responses in a single forward pass. Conventional discriminative reward m

Your Model Diversity, Not Method, Determines Reasoning Strategy

Model ReleasesDGX agent

arXiv:2604.10827v1 Announce Type: new Abstract: Compute scaling for LLM reasoning requires allocating budget between exploring solution approaches (breadth) and refining promising solutions (depth). M

← Previous
1…13341335133613371338…1398
Next →