AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlog
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
Research

Predicting the Emergence of Induction Heads in Language Model Pretraining

DGX agent

arXiv:2511.16893v3 Announce Type: replace Abstract: Specialized attention heads dubbed induction heads (IHs) have been argued to underlie the remarkable in-context learning capabilities of modern lang

researcharxiv-cs-cl
7 Jul 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering

DGX agent

arXiv:2607.04163v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable progress in visual understanding tasks such as image captioning and visual question answ

researcharxiv-cs-ai
7 Jul 2026
Research

Selective Disclosure Watermarking for Large Language Models

DGX agent

arXiv:2607.05353v1 Announce Type: cross Abstract: Watermarking methods embed imperceptible and verifiable signals into text generated by large language models (LLMs). Existing approaches include zero-

researcharxiv-cs-ai
7 Jul 2026
Research

SleepBand: Single-Source Domain Generalization for Sleep Staging via Physiologically Structured Spectral Modeling

DGX agent

arXiv:2607.04851v1 Announce Type: cross Abstract: Generalizing sleep staging models to unseen datasets is challenging, and typical domain generalization (DG) methods often rely on multiple source doma

researcharxiv-cs-lg
7 Jul 2026
Industry

Sources: Beijing recently held meetings with Alibaba, ByteDance, Z.ai, and others to discuss restricting overseas access to advanced open and closed AI models (Fanny Potkin/Reuters)

DGX agent

Fanny Potkin / Reuters: Sources: Beijing recently held meetings with Alibaba, ByteDance, Z.ai, and others to discuss restricting overseas access to advanced open and closed AI models — Chinese authori

industrytechmeme
7 Jul 2026
Applications

Sovereign AI strategies of all types are built on the assumption of continuous releases of open weights models that keep pace with the front…

DGX agent

Sovereign AI strategies of all types are built on the assumption of continuous releases of open weights models that keep pace with the frontier, giving cost/privacy/control gains at the expense of onl

applicationsethan-mollick--x
7 Jul 2026
Safety

States are asking for a $1.4 trillion fine of Meta over addicting kids. Basically they are saying Facebook's business model is a result of c…

DGX agent

States are asking for a $1.4 trillion fine of Meta over addicting kids. Basically they are saying Facebook's business model is a result of crime, and all of Mark Zuckerberg's property should be forfei

safetygary-marcus--x
7 Jul 2026
Agents

Strouhal-Aware Model Predictive Control for Efficient Multi-Fin Flapping Locomotion

DGX agent

arXiv:2607.03216v1 Announce Type: new Abstract: Efficient flapping propulsion hinges on operating within a narrow Strouhal number window, a principle nature has converged upon for maximum thrust-to-po

agentsarxiv-cs-ro
7 Jul 2026
Research

Super-fun in-depth discussion 🙏🙏🙏 World models, JEPA vs generative architectures, anti-collapse methods for JEPA SSL including contrastiv…

DGX agent

Super-fun in-depth discussion 🙏🙏🙏 World models, JEPA vs generative architectures, anti-collapse methods for JEPA SSL including contrastive, distillation, and information maximization methods, an expla

researchyann-lecun--x
7 Jul 2026
Agents

the “harness” will never go away, and it’s just as important as the model

DGX agent

the “harness” will never go away, and it’s just as important as the model new post on harness engineering for AI self-improvement: https://lilianweng.github.io/posts/2026-07-04-harness/ It is hard to

agentsyohei-nakajima--x
7 Jul 2026
Industry

The solution to American open-source lagging behind Chinese open-source: @elonmusk & @cursor_ai releasing their model tomorrow in open-sourc…

DGX agent

The solution to American open-source lagging behind Chinese open-source: @elonmusk & @cursor_ai releasing their model tomorrow in open-source! It's that simple and that would be a massive contribution

industryclem-delangue--x
7 Jul 2026
Applications

This is a key reason I don’t expect the flow of frontier open weights models to continue indefinitely, or even for very much longer. https:/…

DGX agent

This is a key reason I don’t expect the flow of frontier open weights models to continue indefinitely, or even for very much longer. https://www.reuters.com/world/beijing-is-looking-curbing-overseas-a

applicationsethan-mollick--x
7 Jul 2026
Research

Transfer Learning in High-dimensional Ising Models

DGX agent

arXiv:2607.03005v1 Announce Type: new Abstract: In high-dimensional Ising model estimation, target sample sizes are often limited, and effectively using auxiliary binary datasets of unknown relevance

researcharxiv-cs-lg
7 Jul 2026
Safety

Uncertainty-Aware Abstention in Large Language Models with Provable Alignment Guarantees

DGX agent

arXiv:2607.04430v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in question answering (QA) systems, yet they may generate hallucinated or misaligned responses wi

safetyarxiv-cs-cl
7 Jul 2026
Research

Using Mechanistic Interpretability to Craft Adversarial Attacks against Large Language Models

DGX agent

arXiv:2503.06269v3 Announce Type: replace-cross Abstract: Traditional white-box methods for creating adversarial perturbations against LLMs typically rely only on gradient computation from the targete

researcharxiv-cs-ai
7 Jul 2026
Research

Vidu S1: A Real-Time Interactive Video Generation Model

DGX agent

arXiv:2607.03118v1 Announce Type: new Abstract: We introduce Vidu S1, a real-time interactive video generation model supporting voice control of digital characters. Users can control video generation

researcharxiv-cs-cv
7 Jul 2026
Research

Vision Token Manipulation Attacks on Cloud-Edge Inference of Large Vision-Language Models

DGX agent

arXiv:2607.02819v1 Announce Type: cross Abstract: Cloud-edge Large Vision-Language Model (LVLM) inference enables efficient deployment by splitting computation between edge devices and cloud servers.

researcharxiv-cs-ai
7 Jul 2026
Safety

VLM-CASE: Vision-Language Model Enabled Context-Adaptive Safety Envelopes for Anticipatory Safe Autonomous Driving

DGX agent

arXiv:2607.05180v1 Announce Type: cross Abstract: Adverse driving conditions, such as bad weather, remain a principal barrier to autonomous driving because they degrade two things at once: what the ve

safetyarxiv-cs-cv
7 Jul 2026
Research

We sat down with @ylecun for a 1h30 discussion on world models @medjawii and JB Kempf of @videolan/@FFmpeg. It's probably the deepest techni…

DGX agent

We sat down with @ylecun for a 1h30 discussion on world models @medjawii and JB Kempf of @videolan/@FFmpeg. It's probably the deepest technical discussion on the subject ever recorded. And we added En

researchyann-lecun--x
7 Jul 2026
Tutorials

WeightCLIP: Aligning Datasets and Models for Weight Space Learning

DGX agent

arXiv:2607.03551v1 Announce Type: new Abstract: Weight space learning aims to learn representations of neural network (NN) weights, enabling different downstream tasks. Existing approaches show promis

tutorialsarxiv-cs-lg
7 Jul 2026
Agents

Hy3, the new 295B MoE model from @TencentHunyuan, is now free in Nous Portal for the next two weeks! It is focused on cost-effective agentic…

DGX agent

Hy3, the new 295B MoE model from @TencentHunyuan, is now free in Nous Portal for the next two weeks! It is focused on cost-effective agentic use, and particularly strong on coding, tool-calling reliab

agentsnous-research--x
6 Jul 2026
Research

Modeling the chemistry of fusion reactor material

DGX agent

IBM Research explores the application of quantum computing to model the chemical behavior of molten salts used in fusion reactor materials and cooling systems. The work addresses the computational com

researchibm-research
6 Jul 2026
Industry

my friends are all feeling extremely productive and also extremely drained with the latest coding models. this makes me feel like something …

DGX agent

my friends are all feeling extremely productive and also extremely drained with the latest coding models. this makes me feel like something is wrong, and also that there might be a big opportunity. do

industrydavid-holz--x
6 Jul 2026
Local Ai

OpenClaw landed on @huggingface local apps 🦞🤝🤗 1. Pick any GGUF/MLX model on hf 2. Copy the openclaw onboard setup 3. Volla you've got a …

DGX agent

OpenClaw landed on @huggingface local apps 🦞🤝🤗 1. Pick any GGUF/MLX model on hf 2. Copy the openclaw onboard setup 3. Volla you've got a tool-calling agent running fully local. no cloud, no keys, no o

local-aiclem-delangue--x
6 Jul 2026
Agents

We're teaming up with Hugging Face to make open agent traces the fuel for the next generation of open coding models. You can now upload @Dro…

DGX agent

Hugging Face is partnering to make open agent traces available as training data for developing next-generation open-source coding models. This collaboration enables developers to upload agent traces,

agentsclem-delangue--x
6 Jul 2026
Research

The funding headline is only half the story. Yann LeCun told BBC that large language models are not a path to human-like or animal-like inte…

DGX agent

The funding headline is only half the story. Yann LeCun told BBC that large language models are not a path to human-like or animal-like intelligence... #AI https://stechtimes.com/en/article/yann-lecun

researchyann-lecun--x
5 Jul 2026
Research

Warm take: Your world model should never stop learning Introducing AdaJEPA, an adaptive WM that plans, acts, and adapts in a closed loop. Ev…

DGX agent

Warm take: Your world model should never stop learning Introducing AdaJEPA, an adaptive WM that plans, acts, and adapts in a closed loop. Every action leads to a new observation, and every transition

researchyann-lecun--x
5 Jul 2026
Agents

Been going deep in learning @Langchain ‘create_deep_agent’ Turns out create_deep_agent = create_agent (the model <-> tools loop) & a default…

DGX agent

Been going deep in learning @Langchain ‘create_deep_agent’ Turns out create_deep_agent = create_agent (the model <-> tools loop) & a default middleware stack++ 2 nodes do the work, the rest are middle

agentsharrison-chase--x
4 Jul 2026
Industry

Ford is done. GM is done. Honda is done. Toyota is dying. Now FOUR of my employees have bought Tesla Model Y vehicles. They let me drive one…

DGX agent

Ford is done. GM is done. Honda is done. Toyota is dying. Now FOUR of my employees have bought Tesla Model Y vehicles. They let me drive one. I was blown away... again. Full-Self Driving (FSD) is high

industryelon-musk--x
4 Jul 2026
Tutorials

Happy July 4th everyone 🇺🇸 250 years ago, our founding fathers did not have access to frontier models. (Alexander Hamilton wrote 51 of the…

DGX agent

Happy July 4th everyone 🇺🇸 250 years ago, our founding fathers did not have access to frontier models. (Alexander Hamilton wrote 51 of the 85 Federalist papers. Imagine how many he would've written if

tutorialsjerry-liu--x
4 Jul 2026
Agents

Hermes Agent is built for sovereignty and constructing your AI stack how you want and need it to be. No vendor lockins, no model limitations…

DGX agent

Hermes Agent is built for sovereignty and constructing your AI stack how you want and need it to be. No vendor lockins, no model limitations, and most importantly, your IP is built through the self im

agentsnous-research--x
4 Jul 2026
Tools

Open-source models give you complete control, customization, and ownership over your data. Companies are moving fast on this. @vipulved on @…

DGX agent

Open-source AI models provide users with full control, customization capabilities, and data ownership, addressing privacy and autonomy concerns compared to proprietary alternatives. The post highlight

toolstogether-ai--x
4 Jul 2026
Applications

A big part of how the team at Runway is able to build and train the models we do and serve the inference demand we have is that we've built …

DGX agent

A big part of how the team at Runway is able to build and train the models we do and serve the inference demand we have is that we've built incredibly robust research infrastructure and tooling for al

applicationscristobal-valenzuela--x
3 Jul 2026
Agents

A Practice Auditing Framework for Large Language Model Use: Collective Empiricism, Pseudo-Rational Cognition, and Governance of AI-Generated Content

DGX agent

arXiv:2607.01248v1 Announce Type: cross Abstract: Large language models are increasingly used for knowledge acquisition, code generation, academic writing, and agent-based automation. In these setting

agentsarxiv-cs-ai
3 Jul 2026
Applications

At Cohere we deploy our models directly to our customers, instead of them sending data to us. It makes our job harder, but their business mo…

DGX agent

At Cohere we deploy our models directly to our customers, instead of them sending data to us. It makes our job harder, but their business more secure: “When you’re using a consumer app, they are using

applicationscohere--x
3 Jul 2026
Research

BamiBERT: A New BERT-based Language Model for Vietnamese

DGX agent

arXiv:2607.02259v1 Announce Type: new Abstract: In this paper, we introduce BamiBERT, a new BERT-based pre-trained language model for Vietnamese that addresses key limitations of PhoBERT -- the curren

researcharxiv-cs-cl
3 Jul 2026
Research

CreativityNeuro: Steering Language Model Weights to Improve Divergent Thinking and Reduce Mode Collapse

DGX agent

arXiv:2607.01433v1 Announce Type: new Abstract: Divergent thinking is a crucial aspect of creativity, yet large language models (LLMs) tend to consistently generate similar responses to open-ended que

researcharxiv-cs-ai
3 Jul 2026
Safety

DriveVLM-RL: Neuroscience-Inspired Reinforcement Learning with Vision-Language Models for Safe and Deployable Autonomous Driving

DGX agent

arXiv:2603.18315v2 Announce Type: replace-cross Abstract: Traditional reinforcement learning (RL) methods rely on manually engineered rewards or sparse collision signals, which fail to capture the ric

safetyarxiv-cs-ai
3 Jul 2026
Research

Estimating Individualized Treatment Effects in Acute Ischemic Stroke with Causal Transformation Models (TRAM-DAG): A Multi-Centre Observational Study with External RCT Validation

DGX agent

arXiv:2606.12623v3 Announce Type: replace-cross Abstract: Personalized medicine in acute ischemic stroke requires moving beyond average treatment effects (ATE) to individualized treatment effect (ITE)

researcharxiv-cs-lg
3 Jul 2026
Industry

India's IT secretary said the country is investigating a data breach at Apple supplier Tata, which exposed files that included photos of iPhone 18 Pro models (Reuters)

DGX agent

Reuters: India's IT secretary said the country is investigating a data breach at Apple supplier Tata, which exposed files that included photos of iPhone 18 Pro models — India is investigating a data b

industrytechmeme
3 Jul 2026
Applications

Large language models reshape the language of science

DGX agent

arXiv:2504.12317v2 Announce Type: replace Abstract: Scientific language is a central infrastructure of knowledge production, but it remains unclear whether large language models (LLMs) are altering no

applicationsarxiv-cs-cl
3 Jul 2026
Agents

Leveraging Metamemory Agent for Enhanced Data-Free Code Generation in Large Language Models

DGX agent

arXiv:2501.07892v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong performance in automated code generation, with few-shot prompting widely used for its simplicit

agentsarxiv-cs-ai
3 Jul 2026
Safety

Optimizing Visual Generative Models via Distribution-wise Rewards

DGX agent

arXiv:2607.02291v1 Announce Type: new Abstract: Conventional reinforcement learning strategies for visual generation typically employ sample-wise reward functions, yet this practice frequently results

safetyarxiv-cs-lg
3 Jul 2026
Local Ai

Procedural Memory Distillation: Online Reflection for Self-Improving Language Models

DGX agent

arXiv:2607.01480v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR), along with recent selfdistillation variants such as SDPO, evaluates each rollout against a verifi

local-aiarxiv-cs-ai
3 Jul 2026
Safety

Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model

DGX agent

arXiv:2607.01595v1 Announce Type: new Abstract: As the scale and complexity of cloud-based AI systems continue to escalate, ensuring service reliability through rapid fault detection and adaptive reco

safetyarxiv-cs-ai
3 Jul 2026
Research

Scaling Latent Reasoning via Looped Language Models

DGX agent

arXiv:2510.25741v5 Announce Type: replace Abstract: Modern LLMs are trained to 'think' primarily via explicit text generation, such as chain-of-thought (CoT), which defers reasoning to post-training a

researcharxiv-cs-cl
3 Jul 2026
Research

Towards Cellular-Scale Interpretability in Pathology Foundation Models for Biomarker Assessment

DGX agent

arXiv:2511.05150v2 Announce Type: replace-cross Abstract: Molecular biomarker testing in pathology is often costly and tissue-consuming, limiting scalable clinical deployment. Artificial intelligence

researcharxiv-cs-ai
3 Jul 2026
Safety

Transport Discrepancy as a Reliability Signal for Vision-Language-Action Models

DGX agent

arXiv:2512.01715v2 Announce Type: replace Abstract: Vision-language-action (VLA) models that generate continuous action chunks via flow matching lack an internal signal for judging whether a given pre

safetyarxiv-cs-ro
3 Jul 2026
← Previous
1…262263264265266…1293
Next →