AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,483 results
25 May 2026

Linear Regression with Unknown Truncation Beyond Gaussian Features

ResearchDGX agent

arXiv:2602.12534v2 Announce Type: replace-cross Abstract: In truncated linear regression, samples (x,y) are shown only when the outcome y falls inside a certain survival set S^star and the goal is to

Lipschitz Optimization for Formal Verification of Homographies

Model ReleasesDGX agent

arXiv:2605.23203v1 Announce Type: cross Abstract: The adoption of vision neural networks in regulated industries requires formal robustness guarantees, especially in safety-critical domains such as he

Listen to my conversation with @jerryjliu0 of @llama_index on Spotify: https://bit.ly/3v1R0Tu Apple: https://bit.ly/4bTCwpD Youtube: https:/…

TutorialsDGX agent

Listen to my conversation with @jerryjliu0 of @llama_index on Spotify: https://bit.ly/3v1R0Tu Apple: https://bit.ly/4bTCwpD Youtube: https://bit.ly/3uXthnv LinkedIn: http://bit.ly/3Xs8GQP Website: htt

Content type
AllBlogX PostPaperYouTubeRedditGitHub

LLAMA LIMA: A Living Meta-Analysis on the Effects of Generative AI on Learning Mathematics

Model ReleasesDGX agent

arXiv:2601.18685v3 Announce Type: replace-cross Abstract: The capabilities of generative AI in mathematics education are rapidly evolving, posing significant challenges for research to keep pace. Rese

LlamaParse now parses HEIC files natively 🎉 . HEIC is Apple's default image format, so it shows up all over enterprise file systems. Photos…

ApplicationsDGX agent

LlamaParse now parses HEIC files natively 🎉 . HEIC is Apple's default image format, so it shows up all over enterprise file systems. Photos of whiteboards, scanned docs, receipts snapped on an iPhone.

LLM Code Smells: A Taxonomy and Detection Approach

ResearchDGX agent

arXiv:2605.22976v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly integrated into software systems for diverse purposes, due to their versatility, flexibility, and abilit

LLM-driven design of physics-constrained constitutive models: two agents are better than one

Model ReleasesDGX agent

arXiv:2605.23754v1 Announce Type: new Abstract: Developing constitutive models that capture how materials deform under load traditionally requires years of specialized expertise in continuum mechanics

LLM Sparsity Prior for Robust Feature Selection

ResearchDGX agent

arXiv:2605.23102v1 Announce Type: cross Abstract: Large language models (LLMs) offer a scalable mechanism to elicit domain-informed prior information for high-dimensional variable selection. However,

LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws

ResearchDGX agent

arXiv:2605.23901v1 Announce Type: cross Abstract: Existing scaling laws for Large Language Models (LLMs), predominantly monotonic power laws, fail to explain emerging non-monotonic phenomena such as c

Low-Cost Hard-Label Adversarial Attack with Theoretical Foundations

ResearchDGX agent

arXiv:2601.14300v3 Announce Type: replace Abstract: Hard-label black-box attacks, relying solely on top-1 predictions, represent one of the most challenging yet practically threat models. Despite rece

LQ-rPPG: A Label-Quantized Coarse-to-Fine Learning Framework for Remote Physiological Measurement

Model ReleasesDGX agent

arXiv:2605.23174v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) enables non-contact measurement of physiological signals from facial videos, offering strong potential for remote hea

Machine learning applied to emerald gemstone grading: framework proposal and creation of a public dataset

ResearchDGX agent

arXiv:2605.23777v1 Announce Type: new Abstract: The grading of gemstones is currently a manual procedure performed by gemologists. A popular approach uses reference stones, where those are visually in

MadEvolve: Evolutionary Optimization of Trading Systems with Large Language Models

Model ReleasesDGX agent

arXiv:2605.23007v1 Announce Type: cross Abstract: We explore the application of LLM-driven algorithm optimization to several common tasks in quantitative finance. MadEvolve, a general-purpose algorith

Major difference in my mind: - an engineer, given a problem, invents and tries multiple solutions and stops when the solution is good enough…

ResearchDGX agent

Major difference in my mind: - an engineer, given a problem, invents and tries multiple solutions and stops when the solution is good enough. The goal is product innovation and shipping. - a scientist

MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems

SafetyDGX agent

arXiv:2602.04431v2 Announce Type: replace Abstract: LLM-based multi-agent systems have demonstrated impressive capabilities, but they also introduce significant safety risks when individual agents fai

MapGCLR: Geospatial Contrastive Learning of Representations for Online Vectorized HD Map Construction

AgentsDGX agent

arXiv:2603.10688v2 Announce Type: replace-cross Abstract: Autonomous vehicles rely on map information to understand the world around them. However, the creation and maintenance of offline high-definit

MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination

AgentsDGX agent

arXiv:2605.22949v1 Announce Type: new Abstract: Foundation model agents increasingly operate in multi-agent deployments where a coordinator must decide which agent's response to trust. The standard ap

MARS: Magnitude-Aware Rank Statistics

ResearchDGX agent

arXiv:2605.23563v1 Announce Type: new Abstract: Comprehensive evaluation of machine learning models is the key to make sure that they perform as robustly and consistently as desired. In order to summa

MAS-Orchestra: Understanding and Improving Multi-Agent Reasoning Through Holistic Orchestration and Controlled Benchmarks

Model ReleasesDGX agent

arXiv:2601.14652v5 Announce Type: replace Abstract: While multi-agent systems (MAS) promise elevated intelligence through coordination of agents, current approaches to automatic MAS design under-deliv

Massachusetts formally recognizes the App Drivers Union, which says it represents ~70,000 workers and is the first state-certified rideshare union in the US (Bryan Hecht/The Boston Globe)

IndustryDGX agent

Bryan Hecht / The Boston Globe: Massachusetts formally recognizes the App Drivers Union, which says it represents ~70,000 workers and is the first state-certified rideshare union in the US — Uber and

MDS-DETR: DETR with Masked Duplicate Suppressor

ResearchDGX agent

arXiv:2605.23507v1 Announce Type: new Abstract: The DEtection TRansformer (DETR) is a powerful end-to-end object detector, yet its one-to-one matching strategy suffers from slow convergence and low re

MedExpMem: Adapting Experience Memory for Differential Diagnosis

Model ReleasesDGX agent

arXiv:2605.22872v1 Announce Type: cross Abstract: Experienced physicians develop diagnostic expertise through clinical practice, acquiring not only disease knowledge but also the ability to differenti

Mediative Fuzzy Logic: From Type-1 Foundations to Type-2, Type-3 and Quantum Extensions

SafetyDGX agent

arXiv:2605.22900v1 Announce Type: new Abstract: Mediative Fuzzy Logic was conceived as a practical scheme for reconciling hesitant or conflicting assessments in fuzzy control and decision-making. Howe

MedSAE: Dissecting MedCLIP Representations with Sparse Autoencoders

ApplicationsDGX agent

arXiv:2510.26411v2 Announce Type: replace Abstract: Artificial intelligence in healthcare requires models that are accurate and interpretable. We advance mechanistic interpretability in medical vision

MELT: A Behavioral Trace Dataset for High-Risk Memecoin Launch Detection

Model ReleasesDGX agent

arXiv:2602.13480v2 Announce Type: cross Abstract: Launchpads have become the dominant mechanism for issuing memecoins, exposing investors to a new class of high-risk launches that existing rug-pull de

MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structural Anomaly Detection

AgentsDGX agent

arXiv:2605.23723v1 Announce Type: new Abstract: Large language model agents increasingly rely on persistent memory to store past interactions, retrieve relevant demonstrations, and improve long-horizo

Memorial Day. Today we honor those who gave everything to defend our democracy. Democracy isn't guaranteed; it's a precious inheritance that…

TutorialsDGX agent

Andrew Ng posted a Memorial Day tribute emphasizing that democracy requires sacrifice and cannot be taken for granted, describing it as a 'precious inheritance' that demands defense. The post honors t

Memorization Dynamics of Fill-in-the-Middle Pretraining

Model ReleasesDGX agent

arXiv:2605.22981v1 Announce Type: cross Abstract: Fill-in-the-middle (FIM) is a pretraining objective widely used to equip causal language models with infilling ability, yet its effect on verbatim mem

Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals

ResearchDGX agent

arXiv:2605.23384v1 Announce Type: cross Abstract: Recent RL methods have substantially improved the reasoning abilities of LLMs. Existing reward designs mainly follow two paradigms: (1) Reinforcement

Metadata Predictability Is Not Evidence Dependence: An Intervention-Based Audit for Weak-Label Benchmarks

Model ReleasesDGX agent

arXiv:2605.23701v1 Announce Type: new Abstract: We study a protocol-level test for weak-label benchmarks: whether benchmark outputs change when the provided evidence is intervened on. Metadata-only sh

Microsoft economist's hot take: Let it burn first

IndustryDGX agent

This post discusses how OpenAI is burning 17 billion annually despite 20+ billion in revenue because unit economics for AI inference are fundamentally broken, with the company losing more money on que

Millimeter-wave Imaging for Anthropometric Body Measurement

SafetyDGX agent

arXiv:2605.23064v1 Announce Type: new Abstract: Body shape and circumferences are clinically informative biomarkers for risk stratification, including measures such as waist to hip ratio, limb and tru

MiniCPM5-1B is an impressive release in the 1B class! @OpenBMB https://huggingface.co/collections/openbmb/minicpm5 ✨ 1B - Apache 2.0 ✨ Hybri…

HardwareDGX agent

MiniCPM5-1B is an impressive release in the 1B class! @OpenBMB https://huggingface.co/collections/openbmb/minicpm5 ✨ 1B - Apache 2.0 ✨ Hybrid reasoning with Think / No-Think modes ✨ 128K context ✨ Run

MirrorCheck: Efficient Adversarial Defense for Vision-Language Models

ResearchDGX agent

arXiv:2406.09250v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly susceptible to sophisticated adversarial attacks, including adaptive strategies specifically de

Mitigating Object Hallucinations via Sentence-Level Early Intervention

ResearchDGX agent

arXiv:2507.12455v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have revolutionized cross-modal understanding but continue to struggle with hallucinations - fabricated con

Model Collapse as Cultural Evolution

Model ReleasesDGX agent

arXiv:2605.23054v1 Announce Type: cross Abstract: Model collapse, the progressive degradation of LLMs trained on their own outputs, has been characterized statistically but lacks a linguistic explanat

ModeSwitch-LLM: A Lightweight Phase-Aware Controller for Cross-Mode LLM Inference on a Single GPU

Model ReleasesDGX agent

arXiv:2605.23057v1 Announce Type: cross Abstract: ModeSwitch-LLM is a lightweight request-boundary controller for improving single-GPU large language model inference efficiency by routing each request

Moonwalk: Inverse-Forward Differentiation

Model ReleasesDGX agent

arXiv:2402.14212v4 Announce Type: replace-cross Abstract: Backpropagation's main limitation is its need to store intermediate activations (residuals) during the forward pass, which restricts the depth

More details on the Datasette blog: https://datasette.io/blog/2026/jump-menu/

ToolsDGX agent

Datasette, an open-source tool for exploring and publishing data, released a blog post detailing a new jump menu feature for 2026. The jump menu likely provides improved navigation and quick access to

More than 5,500 GitHub repositories were infected with malware in a supply chain attack, dubbed Megalodon, on May 18 that relies on automated commits (Ionut Arghire/SecurityWeek)

IndustryDGX agent

Ionut Arghire / SecurityWeek: More than 5,500 GitHub repositories were infected with malware in a supply chain attack, dubbed Megalodon, on May 18 that relies on automated commits — Fake automated com

Move on Muon : A Hamiltonian probability gradient flow perspective of Muon optimizer

Model ReleasesDGX agent

arXiv:2605.23871v1 Announce Type: cross Abstract: We develop a gradient flow on the space of probability measures defined on matrix-valued parameters induced by regularized Muon, an analytically smoot

MuellerPT: Decomposition Driven Pretraining for Dense Learning in Mueller Polarimetry

ResearchDGX agent

arXiv:2605.23840v1 Announce Type: new Abstract: Mueller matrix imaging provides rich, physically meaningful contrast for biomedical tissue analysis, but supervised learning is hindered by scarce dense

Multi-Floor Exploration for Ground Robots via an Incremental Reachable Graph and Structural Priors

AgentsDGX agent

arXiv:2605.23350v1 Announce Type: new Abstract: Autonomous exploration of multi-floor buildings remains challenging for ground robots because conventional 2D and 2.5D maps cannot represent overlapping

Multi-Gate Residuals

ResearchDGX agent

arXiv:2605.23259v1 Announce Type: cross Abstract: While Attention Residuals has shown some effectiveness in addressing the widespread issue of unbounded activation growth across deep residual layers,

Multi-SpatialMLLM: Multi-Frame Spatial Understanding with Multi-Modal Large Language Models

Model ReleasesDGX agent

arXiv:2505.17015v2 Announce Type: replace-cross Abstract: Multi-modal large language models (MLLMs) have rapidly advanced in visual tasks, yet their spatial understanding remains limited to single ima

Multilingual Knowledge Transfer under Data Constraints via Lexical Interventions

ResearchDGX agent

arXiv:2605.23885v1 Announce Type: new Abstract: Cross-lingual knowledge transfer is critical for building high-performing multilingual language models for languages with insufficient training data. Wh

Multilingual Steering by Design: Multilingual Sparse Autoencoders and Principled Layer Selection

Model ReleasesDGX agent

arXiv:2605.23036v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) enable feature-level mechanistic interpretability and activation steering in large language models (LLMs), but SAE-based lang

Multimodal Crystal Flow: Any-to-Any Modality Generation for Unified Crystal Modeling

ResearchDGX agent

arXiv:2602.20210v2 Announce Type: replace-cross Abstract: Crystal modeling spans a family of conditional and unconditional generation tasks, including crystal structure prediction (CSP) and de novo ge

Multimodal Distribution Matching for Vision-Language Dataset Distillation

SafetyDGX agent

arXiv:2605.23482v1 Announce Type: cross Abstract: Dataset distillation compresses large training sets into compact synthetic datasets while preserving downstream performance. As modern systems increas

MUSEKG: A Knowledge Graph Over Museum Collections

ResearchDGX agent

arXiv:2511.16014v2 Announce Type: replace Abstract: Digitisation in the cultural heritage sector has produced large but fragmented repositories of museum collection data, spanning structured catalogue

Musk hates Altman because Altman deceived him. Altman hates Musk because Musk is an egomaniac. LeCun (who is an equally big egomaniac) hates…

SafetyDGX agent

Musk hates Altman because Altman deceived him. Altman hates Musk because Musk is an egomaniac. LeCun (who is an equally big egomaniac) hates Musk because Musk is a jerk. Musk hates LeCun because LeCun

My favorite prompt: a) make a plan for <task> b) orchestrate and launch sub-agents to execute the plan c) validate the results from the sub-…

IndustryDGX agent

My favorite prompt: a) make a plan for <task> b) orchestrate and launch sub-agents to execute the plan c) validate the results from the sub-agents d) repeat b and c until you finish the plan Grok Buil

Naturalistic measure of social norms alignment

SafetyDGX agent

arXiv:2605.23420v1 Announce Type: new Abstract: Social norms reflect shared expectations on acceptable behavior. Measuring social norms alignment remains challenging, with existing approaches typicall

Near-Optimal Private Linear Regression via Iterative Hessian Mixing

ResearchDGX agent

arXiv:2601.07545v2 Announce Type: replace Abstract: We study differentially private ordinary least squares (DP-OLS) with bounded data (X,Y) via sketching-based mechanisms. While Gaussian sketching app

NeuralBoneReg: An Instance-Specific Label-Free Point Cloud-Based Method for Multi-Modal Bone Surface Registration

SafetyDGX agent

arXiv:2511.14286v2 Announce Type: replace Abstract: In computer- and robot-assisted orthopedic surgery (CAOS), patient-specific surgical plans derived from preoperative imaging define target locations

NeuroNL2LTL: A Neurosymbolic Framework for Natural Language Translation of Linear Temporal Logic

SafetyDGX agent

arXiv:2605.22874v1 Announce Type: new Abstract: Effectively translating between natural language (NL) and formal logics like Linear Temporal Logic (LTL) requires expertise that limits formal verificat

NeuroWeaver: An Autonomous Evolutionary Agent for Exploring the Programmatic Space of EEG Analysis Pipelines

AgentsDGX agent

arXiv:2602.13473v2 Announce Type: replace Abstract: Although foundation models have demonstrated remarkable success in general domains, the application of these models to electroencephalography (EEG)

New paper from Microsoft on Self-Evolving Agent Skills

AgentsDGX agent

New paper from Microsoft on Self-Evolving Agent Skills New research from Microsoft Research I see a lot of AI engineers handwriting agent skill docs and hope they generalize. Probably not optimal. Thi

Next-Latent Prediction Transformers Learn Compact World Models

SafetyDGX agent

arXiv:2511.05963v2 Announce Type: replace Abstract: Transformers replace recurrence with a memory that grows with sequence length and self-attention that enables ad-hoc lookups over past tokens. Conse

nice write up from the HuggingFace folks aggregating works on defining agents, harnesses, environments, RL, etc. The more we can roughly hav…

AgentsDGX agent

nice write up from the HuggingFace folks aggregating works on defining agents, harnesses, environments, RL, etc. The more we can roughly have a shared vocabulary the better…I still find it confusing (

← Previous
1…838839840841842…1475
Next →