AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,375 results
30 Jun 2026

The gap in autonomous agentic loops that gets ignored: agents can plan and call APIs but can't acquire tools they don't have access to. x402…

AgentsDGX agent

The gap in autonomous agentic loops that gets ignored: agents can plan and call APIs but can't acquire tools they don't have access to. x402 + Apify's 20,000+ Actors is a concrete fix for that. Worth

The Heterogeneous Safety Impacts of Benign Multilingual Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.28843v1 Announce Type: cross Abstract: Fine-tuning a large language model is a ubiquitous method for enhancing its capability on a specific downstream task. However, prior work has shown th

The Hidden Cost of Resampling: How Imbalance Correction Degrades Probability Calibration in Tree Ensembles

TutorialsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.29720v1 Announce Type: cross Abstract: Resampling methods such as SMOTE and random under/over-sampling are standard tools for class-imbalanced classification, almost always evaluated by min

The Hidden Cost of Structured Generation in LLMs: Draft-Conditioned Constrained Decoding

Model ReleasesDGX agent

arXiv:2603.03305v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to generate executable outputs, JSON objects, and API calls, where a single syntax error ca

The Human Creativity Benchmark

Model ReleasesDGX agent

arXiv:2606.30561v1 Announce Type: new Abstract: Modern AI evaluation frameworks treat evaluator disagreement as noise to be resolved. In creative domains, professional disagreement reflects genuine di

The interesting thing about viral narratives is that they don't have to be true to spread. There's now enough evidence pointing to the fact …

IndustryDGX agent

The interesting thing about viral narratives is that they don't have to be true to spread. There's now enough evidence pointing to the fact that AI usage does not kill jobs, but that might not matter

The Interference Gap: Comparing Retrieval Bounds in Human Memory and RAG Systems

Model ReleasesDGX agent

arXiv:2606.28327v1 Announce Type: cross Abstract: How do retrieval bounds compare between human episodic memory and Retrieval-Augmented Generation (RAG) systems under semantic interference? We present

The Joint Effect of Quantization and Sampling Temperature on LLM Safety Alignment: A Factorial Analysis

SafetyDGX agent

arXiv:2606.29581v1 Announce Type: cross Abstract: Modern LLM deployments routinely compress models and raise sampling temperature to reduce cost, latency, or repetition, yet safety evaluations usually

𝗚𝗟𝗠-𝟱.𝟮 (the latest open weights model) is having an Enterprise moment, and it is not an exaggeration.🚀 🔥 We have been impressed by h…

Model ReleasesDGX agent

𝗚𝗟𝗠-𝟱.𝟮 (the latest open weights model) is having an Enterprise moment, and it is not an exaggeration.🚀 🔥 We have been impressed by how strongly GLM-5.2 is pushing long-horizon performance .. not just

The Many-Body Problem of the Data Centre

ResearchDGX agent

arXiv:2606.30206v1 Announce Type: new Abstract: Modern Artificial Intelligence is often framed as limited by its own disembodiment, as if giving it a body would unlock its true potential. We argue to

The media bias against Elon is absolutely insane They don’t report on him They hunt for ways to make him look evil Is it a lie? Doesn’t matt…

SafetyDGX agent

The media bias against Elon is absolutely insane They don’t report on him They hunt for ways to make him look evil Is it a lie? Doesn’t matter Is there proof? Still doesn’t matter They will lie, twist

The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.29526v1 Announce Type: new Abstract: Reinforcement learning (RL) has gained growing attention in large language model (LLM) post-training, yet RL training remains fragile and can suffer fro

The most important weird thing about LLMs is that they are so general. A bigger LLM that is better at coding is also better at ideation & et…

ApplicationsDGX agent

The most important weird thing about LLMs is that they are so general. A bigger LLM that is better at coding is also better at ideation & ethical advice & medicine & math. This isn’t true of everythin

The NTNU System at the S&I Challenge 2025 SLA Open Track

Model ReleasesDGX agent

arXiv:2506.05121v3 Announce Type: replace Abstract: A recent line of research on spoken language assessment (SLA) employs neural models such as BERT and wav2vec 2.0 (W2V) to evaluate speaking proficie

The Platonic Defense: Backdoor Defense for Self-Supervised Encoders in the Era of Large Scale Pre-training

ResearchDGX agent

arXiv:2606.29451v1 Announce Type: new Abstract: Self-supervised learning (SSL) pretrained models have become a dominant paradigm for visual representation learning, but they are vulnerable to backdoor

The registrar's function in a hybrid society. AI value chain,smart data and the concept of property

ApplicationsDGX agent

arXiv:2606.28789v1 Announce Type: cross Abstract: Artificial intelligence reaches the land registry not as another tool but as a value chain that turns data into intelligence and intelligence into eco

The Speedup Paradox: Rethinking Inference Speed-Quality Trade-off in Embodied Tasks

ResearchDGX agent

arXiv:2606.28529v1 Announce Type: cross Abstract: Embodied foundation models have recently been widely used to improve robot generalization and task success rates. Previous works apply lossy efficient

The strength of clinical evidence is recoverable from language model representations but not from their stated grades

Local AiDGX agent

arXiv:2606.29034v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly summarize clinical evidence, where a claim's weight depends on how strongly it is supported. Yet these model

The Surprising Effectiveness of Video Diffusion Models for Hand Motion Reconstruction

ResearchDGX agent

arXiv:2606.30308v1 Announce Type: new Abstract: 4D hand motion reconstruction from egocentric video is bottlenecked by clear limitations of existing methods: image-based pipelines depend on a detector

The twilight of the chatbots

ApplicationsDGX agent

This article likely examines the limitations and declining effectiveness of current chatbot technology, exploring why large language models may be reaching performance plateaus and what challenges lie

The Two Genie Game: Adoption and Welfare in Audit-Grounded AI Governance

SafetyDGX agent

arXiv:2606.28710v1 Announce Type: new Abstract: We ask under what conditions an agent with a harm-minimizing policy can displace an approval-seeking (RLHF) agent in a competitive market, and when that

The UK CMA proposes requiring Apple and Google to relax UK developer payment 'steering' rules, and Apple to open NFC; Google says it 'already made the changes' (Sam Tabahriti/Reuters)

IndustryDGX agent

Sam Tabahriti / Reuters: The UK CMA proposes requiring Apple and Google to relax UK developer payment “steering” rules, and Apple to open NFC; Google says it “already made the changes” — Britain's com

The Undecidability of Artificial General Intelligence (AGI) Alignment

SafetyDGX agent

arXiv:2606.28639v1 Announce Type: cross Abstract: This article establishes the foundational mathematical limits of Artificial General Intelligence (AGI) safety, proving that the core barrier is not th

The US going 100% EV by 2040 would save more than 100k lives, study says

IndustryDGX agent

A study from the American Lung Association found that if the U.S. transitions to a 100% zero-emission passenger vehicle fleet by 2050, it would result in 89,300 fewer premature deaths and $978 billion

The Verbose Context Problem in Medical Records

Model ReleasesDGX agent

arXiv:2606.29503v1 Announce Type: cross Abstract: The verbose context problem occurs when structured concepts have token-inefficient textual representations. This bottleneck is acute in population hea

The Voronoi Bottleneck: Capacity-Aware Dense Retrieval for Product Search

ResearchDGX agent

arXiv:2606.28359v1 Announce Type: cross Abstract: Dense embedding retrieval compresses all relevance information into a single inner product, imposing a fundamental geometric limit -- the Voronoi Bott

Theory of Continual Learning Against Data Poisoning Attacks

SafetyDGX agent

arXiv:2606.29841v1 Announce Type: new Abstract: Continual learning (CL), where a model is trained on a sequence of data tasks, is increasingly being adopted across key fields such as large language mo

There is a lot of pride among AI founders today around doing '996.' 9 to 9, 6 days a week. SF is normalizing the 72-hour week to win the AI …

ToolsDGX agent

There is a lot of pride among AI founders today around doing '996.' 9 to 9, 6 days a week. SF is normalizing the 72-hour week to win the AI race. I started Upside to enable a different way of winning.

They are such liars!

IndustryDGX agent

They are such liars! The original frame for the US-AID debate was: 'Evil space Nazi Elon Musk killed 14,000,000 people!!!' *I and others pointed out that this is not possible, given that the global de

They’re letting me keynote at this conference for some reason

ToolsDGX agent

Swyx posted about being invited to deliver a keynote address at a conference, expressing humorous surprise at being selected for this speaking role. The post likely reflects on the unexpected opportun

ThinkProbe: Beyond Accuracy -- Structural Profiling of Open-Ended LLM Reasoning Traces via Non-Generative Thought Graphs

ResearchDGX agent

arXiv:2606.29067v1 Announce Type: new Abstract: We present ThinkProbe, a framework for structural analysis of LLM reasoning traces. ThinkProbe converts each trace into a Thought Graph a directed graph

this is what you look like with low rise pants

Model ReleasesDGX agent

This post likely showcases visual examples or commentary on the aesthetic appearance and fit of low-rise pants, a fashion trend that was particularly popular in the early 2000s and has experienced per

this may be ai generated, but is true - the hardest part of wikis (and memory in general) is the process that condenses raw data into learni…

AgentsDGX agent

this may be ai generated, but is true - the hardest part of wikis (and memory in general) is the process that condenses raw data into learnings/memory @hwchase17 @cognition @FactoryAI @karpathy wiki m

Thrilled to announce the Wearable AI Workshop at ECCV 2026 🎉 that we're organizing with an awesome group of folks across Meta Reality Labs,…

Model ReleasesDGX agent

Thrilled to announce the Wearable AI Workshop at ECCV 2026 🎉 that we're organizing with an awesome group of folks across Meta Reality Labs, AMI Labs, HKUST, Georgia Tech, UCF, and U. of Edinburgh. If

Through Aug 31, Sonnet 5 will use roughly 30% less quota than Sonnet 4.6 in Devin Desktop/CLI. After that, it will use the same quota as Son…

AgentsDGX agent

Claude Sonnet 5 will consume approximately 30% less API quota than Sonnet 4.6 when used through Devin Desktop or CLI until August 31, after which both models will use equivalent quota allocations. Thi

Thunder-KoNUBench: A Corpus-Aligned Benchmark for Korean Negation Understanding

Model ReleasesDGX agent

arXiv:2601.04693v2 Announce Type: replace Abstract: Although negation is known to challenge large language models (LLMs), benchmarks for evaluating negation understanding-especially in Korean-are scar

thx @firecrawl for the fast wifi, just across the street

AgentsDGX agent

Yohei Nakajima posted a brief thank you message to Firecrawl for providing fast WiFi service, mentioning the location was conveniently located just across the street. The post appears to be a casual s

Timesteps of Mamba Align with Human Reading Times

SafetyDGX agent

arXiv:2606.29904v1 Announce Type: new Abstract: This study demonstrates an alignment of per-word processing time in a popular state-space language model Mamba and human readers. In Mamba, the recurren

To Reason or to Fabricate: Reasoning Without Shortcuts via Hint-Anchored Pairwise Aggregation

ResearchDGX agent

arXiv:2606.29481v1 Announce Type: cross Abstract: While reinforcement learning (RL) significantly enhances LLM reasoning, its efficacy is severely undermined by Pre-RL data overlap, where RL datasets

To Tab or Not to Tab: Measuring Critical Engagement in AI Code Completion Tools Using Behavioral Signals and Attention Checks

ResearchDGX agent

arXiv:2606.30549v1 Announce Type: cross Abstract: AI code completion tools, such as Github Copilot, provide students with code suggestions to help them write programs. However, recent qualitative stud

To Use or not to Use Muon: How Simplicity Bias in Optimizers Matters

SafetyDGX agent

arXiv:2603.00742v2 Announce Type: replace Abstract: While Adam has long been the ubiquitous default optimizer for deep neural networks, Muon has recently seen rapid adoption due to its superior traini

Today, we give robots a /skills library that self-evolves and compounds indefinitely! Introducing ASPIRE: a robot solving its 100th task is …

Local AiDGX agent

Today, we give robots a /skills library that self-evolves and compounds indefinitely! Introducing ASPIRE: a robot solving its 100th task is no longer as clueless as solving its first. Coding agents ob

Today’s the day. The inaugural Agent Open is here and it’s gonna be awesome. Arrive early, play some pickleball, meet some people, eat good …

AgentsDGX agent

Today’s the day. The inaugural Agent Open is here and it’s gonna be awesome. Arrive early, play some pickleball, meet some people, eat good food, and enjoy the panel featuring @travers00 @ankrgyl @Sir

Together AI at ICML 2026: frontier research across the full stack

ToolsDGX agent

Together AI presented research at ICML 2026 covering advances across the full technology stack, likely including foundational model improvements, inference optimization, and practical deployment solut

TokenBudgeting: Our Conversations with Enterprises on Token Spend

HardwareDGX agent

TokenBudgeting examines enterprise spending patterns and cost management strategies related to AI token consumption, based on SemiAnalysis's direct conversations with corporate customers. The analysis

Tool-Augmented Spatiotemporal Reasoning for Streamlining Video Question Answering Task

AgentsDGX agent

arXiv:2512.10359v1 Announce Type: cross Abstract: Video Question Answering (VideoQA) task serves as a critical playground for evaluating whether foundation models can effectively perceive, understand,

Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2606.28425v1 Announce Type: cross Abstract: Increasingly autonomous agentic AI systems pose novel multi-agent risks, such as secret collusion via covert communication channels. The natural defen

TopoAgent: An Agentic Framework for Automated Topology Learning in Medical Imaging

AgentsDGX agent

arXiv:2606.29763v1 Announce Type: cross Abstract: Topological data analysis (TDA), particularly persistent homology (PH), captures geometric structural properties in medical images (e.g., connected co

Toward an Energy-Optimized Operation of Data Centers Located in Wind Farms Using Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.30316v1 Announce Type: new Abstract: This paper studies Reinforcement Learning as an online controller for curtailment-aware workload shifting in wind-turbine-integrated high-performance co

Toward Secure and Reliable PDDL Formalization of Large Language Models with Planner-in-the-Loop Feedback

Model ReleasesDGX agent

arXiv:2606.29700v1 Announce Type: new Abstract: Planning often requires symbolic specifications that are both executable and verifiable. For large language models deployed in autonomous or decision-su

Towards Biosignals-Free Autonomous Prosthetic Hand Control via Imitation Learning

AgentsDGX agent

arXiv:2506.08795v2 Announce Type: replace-cross Abstract: Limb loss affects millions globally, impairing physical function and reducing quality of life. Most traditional surface electromyographic (sEM

Towards Complete Causal Explanation with Expert Knowledge

ResearchDGX agent

arXiv:2407.07338v4 Announce Type: replace-cross Abstract: We study the problem of restricting a Markov equivalence class of maximal ancestral graphs (MAGs) to only those MAGs that contain certain edge

Towards Continual Motion-Language Agents: LoRA Variants for Incremental Motion Understanding and Generation

Model ReleasesDGX agent

arXiv:2606.30266v1 Announce Type: cross Abstract: Motion-language agents must possess the bidirectional capability to both understand human movement (motion-to-text, M2T) and generate it from natural

Towards Engineering Scaling Laws with Pretraining Data Composition

ResearchDGX agent

arXiv:2606.19781v2 Announce Type: replace-cross Abstract: Neural scaling laws describe how model performance improves as a power law in compute, model size, and dataset size. While well-established fo

Towards Evaluating Data Priors for Tabular Foundation Models

ResearchDGX agent

arXiv:2606.29241v1 Announce Type: new Abstract: Data-generating priors are a central component of tabular foundation models because they define the task distribution used during pretraining. However,

Towards Generalizable and Evidential Nuclear Magnetic Resonance-Based Molecular Structure Elucidation via Large Language Model Agent

Model ReleasesDGX agent

arXiv:2606.29776v1 Announce Type: cross Abstract: Nuclear Magnetic Resonance (NMR) spectroscopy is the gold standard for molecular structure elucidation, yet interpreting complex spectra for unknown m

Towards Harnessing the Collaborative Power of Large and Small Models for Domain Tasks

ResearchDGX agent

arXiv:2504.17421v2 Announce Type: replace-cross Abstract: Large language models (LMs) offer broad generalization capabilities but require vast amounts of data and computational resources for domain-sp

Towards Improved Anomaly Detection for Cloud Cybersecurity via Graph Neural Networks

ApplicationsDGX agent

arXiv:2606.28923v1 Announce Type: new Abstract: Detecting security threats in an organization's cloud computing environment has become necessary due to the increased reliance on cloud infrastructure.

Towards in-the-wild Egocentric 3D Hand-Object Pose Estimation

Model ReleasesDGX agent

arXiv:2606.30598v1 Announce Type: new Abstract: Estimating accurate 3D hand-object pose from in-the-wild egocentric RGB remains challenging due to severe occlusions and ambiguous contact. Existing lea

Towards Long-Form Spatio-Temporal Video Grounding

Local AiDGX agent

arXiv:2602.23294v2 Announce Type: replace Abstract: In real scenarios, videos can span several minutes or even hours. However, existing research on spatio-temporal video grounding (STVG), given a text

← Previous
1…460461462463464…1473
Next →