AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
25 Jul 2026

Best C++ Local Model? (July 24th 2026 Edition :-P)

Model ReleasesDGX agent

I apologize that this question has been asked in various flavors over time, but I couldn't find anything in the posts before that matches the options I have. I have a PC and a mac, both available in m

24 Jul 2026

Artificial Epanorthosis: Why large language models overuse a classical rhetorical figure, and how to mitigate it

TutorialsDGX agent

arXiv:2607.21498v1 Announce Type: cross Abstract: A rhetorical figure that Cicero and Quintilian catalogued two thousand years ago reappears, systematically, in the text of large language models: epan

DynamicMCPBench: A Trace-Grounded, Effect-Scored Benchmark for LLM Agents over Live MCP Servers

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2607.20531v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly deployed over Model Context Protocol (MCP) servers, yet the benchmarks used to evaluate them score th

ExecuGraph: A Multi-Agent, Execution-Grounded Framework for Reliable Backend Code Synthesis with Large Language Models

Local AiDGX agent

arXiv:2607.20499v1 Announce Type: new Abstract: Large Language Models generate plausible backend code, but a single-pass paradigm provides no guarantee of correctness or runtime reliability. We presen

Multimodal Pretraining for Generalizable EEG Representation Learning

Model ReleasesDGX agent

arXiv:2607.21384v1 Announce Type: new Abstract: Electroencephalography (EEG) models used for epilepsy are often limited to specific datasets and tasks. This limited approach can make it challenging to

Representation Robustness Under Executable Reasoning Constraints in Large Language Models for Mathematical Problem Solving

Local AiDGX agent

arXiv:2607.20520v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evaluated on mathematical problem solving, yet prior work often treats representationally equivalent formu

Response drift across frontier large language models

ResearchDGX agent

arXiv:2607.20454v1 Announce Type: cross Abstract: All frontier large language models (LLMs) exhibit response drift -- producing outputs that deviate from expert-validated references -- yet the magnitu

Stochastic Sampling is Epistemically Shallow: The Dimensionality Gap Between Temperature Variation and Model Diversity in LLMs

ResearchDGX agent

arXiv:2607.20464v1 Announce Type: new Abstract: When a language model gives different answers on repeated runs, does that variation reveal what it does not know? Self-consistency turns the variation i

The demand for GLM-5.2 and other frontier-level open models has been surging on Ollama's cloud. We're adding capacity in anticipation for so…

Local AiDGX agent

The demand for GLM-5.2 and other frontier-level open models has been surging on Ollama's cloud. We're adding capacity in anticipation for some very large models next week! To make sure Ollama's infras

Token-Level Entropy Reveals Demographic Disparities in Large Language Models

SafetyDGX agent

arXiv:2501.19337v5 Announce Type: replace Abstract: A name alone measurably reshapes a language model's next-token distribution before a single token is sampled. We measure full-vocabulary Shannon ent

Toward Generalizable Cognitive Impairment Detection with Speech-Based Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.21496v1 Announce Type: cross Abstract: Cognitive impairment (CI) is a growing public health concern. Early and accurate diagnosis is critical for enabling timely intervention and improving

TwistedMerge: Certified Higher-Order Diagnostics and Abstention for Model Merging

SafetyDGX agent

arXiv:2607.20887v1 Announce Type: cross Abstract: Model merging combines independently trained or fine-tuned models, but pairwise alignability does not imply globally consistent alignment. We formulat

VeriSimpl: Robust Optimization Modeling from Natural Language using Simplification-based Verification

ResearchDGX agent

arXiv:2607.20474v1 Announce Type: new Abstract: Natural language interfaces can greatly benefit the accessibility and usability of optimization modeling, and recent advances in large language models (

which model to use on local 24gb mac mini M4 pro

Local AiDGX agent

So, i have been building some apps that should run on the local every user system, tried gemma4 although its fast and great at reasoning its not as good in instructions following and tool calling. tri

23 Jul 2026

Antigen-specific Antibody Multi-modal Foundation Model for Functional Antibody Design

TutorialsDGX agent

arXiv:2607.20057v1 Announce Type: cross Abstract: Antibodies are essential proteins that play a central role in immune recognition by binding specific antigen molecules. Although recent protein langua

ChannelGuard: Safe Models Do Not Compose into Safe Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2607.19430v1 Announce Type: cross Abstract: Multi-agent LLM applications chain a planner, worker agents, a verifier, and a synthesizer, and every hop between agents is an unmonitored channel thr

Delineate Anything v2: A Global Foundation Model for Field Delineation

Model ReleasesDGX agent

arXiv:2607.19069v1 Announce Type: new Abstract: Accurate agricultural field boundary delineation at large scale is a foundational task for food security, supply chain transparency, and carbon accounti

Enhancing next token prediction based pre-training for jet foundation models

ResearchDGX agent

arXiv:2512.04149v2 Announce Type: replace-cross Abstract: Next token prediction is an attractive pre-training task for jet foundation models, in that it is simulation free and enables excellent genera

GeoTrace: Geometry-Aware Trajectory Token Compression for Video Large Language Models

Local AiDGX agent

arXiv:2607.09080v2 Announce Type: replace Abstract: Although Video Large Language Models (Video LLMs) have shown strong performance in video understanding, their efficiency is still limited by the lar

Koopman Dreamer: Spectrally Constrained Latent Dynamics for Stable World-Model Imagination

Local AiDGX agent

arXiv:2607.19719v1 Announce Type: new Abstract: Latent world models improve sample efficiency in continuous control by optimizing policies over imagined latent trajectories, but common neural transiti

Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters and 5.1B active, it's q…

AgentsDGX agent

Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters and 5.1B active, it's quick to run and built for agent workloads: coding, search, r

PAVXploreRL: Physical-Action-Visual World Model Reinforcement Learning with Action Exploration

SafetyDGX agent

arXiv:2607.16602v2 Announce Type: replace Abstract: Action-conditioned world models are a key component of embodied AI, serving as scalable policy evaluators that reduce reliance on expensive real-wor

Physics-Aware Complex-Valued State Space Model with Scattering-Prior Feature Modulation for PolSAR Image Classification

Model ReleasesDGX agent

arXiv:2607.19787v1 Announce Type: cross Abstract: Polarimetric synthetic aperture radar (PolSAR) image classification is a representative task for physics-aware GeoAI, where land-cover semantics are c

PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference

SafetyDGX agent

arXiv:2607.20327v1 Announce Type: new Abstract: Large language models (LLMs) provide strong reasoning capabilities but are expensive to serve at scale, whereas small language models (SLMs) are cheaper

22 Jul 2026

AI has started building its own successors. It's happening right now. Humans still kick most of it off, but the models are already training …

IndustryDGX agent

AI has started building its own successors. It's happening right now. Humans still kick most of it off, but the models are already training the next models, and each version builds a better version fa

21 Jul 2026

lol so it’s possible that OpenAI’s guardrails prevented huggingface from being able to analyze the logs of attacks from… OpenAI’s model. so …

AgentsDGX agent

lol so it’s possible that OpenAI’s guardrails prevented huggingface from being able to analyze the logs of attacks from… OpenAI’s model. so they had to switch to an open source Chinese model 😵‍💫 @Open

20 Jul 2026

Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss.…

SafetyDGX agent

Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss. We’re sharing what we learned from studying a long-running

16 Jul 2026

Adapting Generalist Vehicle Models for High-Speed MPC Across Terrains

ApplicationsDGX agent

arXiv:2607.13319v1 Announce Type: cross Abstract: High-speed off-road autonomy requires precise closed-loop control for a target vehicle while remaining robust across changing terrains. Recent forward

AIMO Interpretability Challenge

Model ReleasesDGX agent

arXiv:2607.13899v1 Announce Type: new Abstract: We propose the AIMO Interpretability Challenge, a competition on distinguishing robust from spurious reasoning in frontier mathematical language models

Evaluating Vision Foundation Models for Pixel and Object Classification in Microscopy

Model ReleasesDGX agent

arXiv:2603.19802v2 Announce Type: replace Abstract: Deep learning underlies most modern approaches and tools in computer vision, including biomedical imaging. However, for interactive semantic segment

EXPLORE: Exploration with Guided Search for Analog Topology Generation using Language Models

Model ReleasesDGX agent

arXiv:2607.13416v1 Announce Type: new Abstract: Automating analog circuit topology design is essential to reduce the extensive manual effort required to meet increasingly diverse and customized applic

From Surface Forecasting to Observability Forecasting: A Latent World Model for Cloud-Aware EO Monitoring

Model ReleasesDGX agent

arXiv:2607.13651v1 Announce Type: new Abstract: The bottleneck of Earth Observation processing chains is not the arrival of new imagery but whether the surface is actually visible when the image arriv

Improving Molecular Property Prediction in Small Language Models Using Graph-based Tools

AgentsDGX agent

arXiv:2607.13115v1 Announce Type: new Abstract: Small language models (SLMs) have shown promise for zero-shot molecular property prediction from SMILES strings, yet they often suffer from structural b

MetaPerch: Learning from metadata for bioacoustics foundation models

ApplicationsDGX agent

arXiv:2607.14072v1 Announce Type: new Abstract: Bioacoustic foundation models rely on large-scale citizen science platforms like Xeno-Canto for geographically and ecologically diverse data. Recent wor

TCAM-Diff: Triplane-Aware Cross-Attention Medical Diffusion Model

TutorialsDGX agent

arXiv:2607.13812v1 Announce Type: cross Abstract: We introduce TCAM-Diff, a novel 3D medical image generation model that reduces the memory requirements to encode and generate high-resolution 3D data.

15 Jul 2026

Comparing Semantic Navigation in Humans and Large Language Models using Natural Language Processing

Model ReleasesDGX agent

arXiv:2607.12195v1 Announce Type: cross Abstract: Semantic memory retrieval can be conceptualized as navigation through conceptual space. We compared semantic search dynamics between humans and three

Efficient Conformal Prediction for Regression Models under Label Noise

ResearchDGX agent

arXiv:2509.15120v2 Announce Type: replace Abstract: In high-stakes scenarios, such as medical imaging applications, it is critical to equip the predictions of a regression model with reliable confiden

Excited for our first general model Inkling -- open weights, 975B, natively multimodal (text, image, audio). Available on Tinker, HuggingFac…

ResearchDGX agent

Excited for our first general model Inkling -- open weights, 975B, natively multimodal (text, image, audio). Available on Tinker, HuggingFace and partners. It is yours to personalize and use openly. I

From Many to Meaningful: Feature-Guided Zero-Shot Chronic Kidney Disease Screening Using Large Language Models

Model ReleasesDGX agent

arXiv:2607.12260v1 Announce Type: new Abstract: Early screening of chronic kidney disease (CKD) is essential for preventing irreversible progression; however, many machine learning (ML)-based screenin

From Observation to Insight: Mechanistic World Models and the Quest for Autonomous Discovery

AgentsDGX agent

arXiv:2607.12474v1 Announce Type: new Abstract: Recent advances in foundation models have transformed AI for Science, enabling remarkably accurate predictive performance across domains ranging from pr

Inkling is our first open model from @thinkymachines and is now available on Tinker! Check out these quotes from Tinker customers on their e…

AgentsDGX agent

Inkling is our first open model from @thinkymachines and is now available on Tinker! Check out these quotes from Tinker customers on their experience with Inkling: @_Mantic_AI: 'Not only does Inkling

MaxSAT-Based Feedback for Guiding Vision-Language Models in Sudoku

TutorialsDGX agent

arXiv:2607.12711v1 Announce Type: new Abstract: Vision--Language Models (VLMs) have recently demonstrated promising performance on structured visual reasoning tasks, including grid-based puzzles. Howe

Model-Based Diffusion Optimal Control for Multi-Robot Motion Planning

SafetyDGX agent

arXiv:2607.12423v1 Announce Type: new Abstract: Multi-Robot Motion Planning in continuous environments, where robots must generate dynamically feasible, collision-free trajectories, is challenging due

PluRel: Synthetic Data unlocks Scaling Laws for Relational Foundation Models

TutorialsDGX agent

arXiv:2602.04029v2 Announce Type: replace-cross Abstract: Relational Foundation Models (RFMs) facilitate data-driven decision-making by learning from complex multi-table databases. However, the divers

Resist and Update: Counterfactual Report Coordinates for Incentive-Compatible LLMs

Model ReleasesDGX agent

arXiv:2607.12985v1 Announce Type: new Abstract: Aligned language models routinely misreport under non-evidential incentive pressure: they agree with a confident user or overstate certainty even when t

The Capacity of Thought: Benchmarking Llama 3.2 in Semantic fMRI Neural Language Decoding and Improving the Huth Encoding-Model Baseline

Model ReleasesDGX agent

arXiv:2607.12079v1 Announce Type: new Abstract: Decoding continuous language from fMRI signals remains a core challenge in non-invasive brain-computer interface research. We present two complementary

14 Jul 2026

Huge if true! We are talking about a 27B multimodal model that runs locally on a phone. That's wild! Bonsai 27B reaches up to 163 tok/s in 1…

Local AiDGX agent

Huge if true! We are talking about a 27B multimodal model that runs locally on a phone. That's wild! Bonsai 27B reaches up to 163 tok/s in 1-bit and 134 tok/s in Ternary on an NVIDIA GeForce RTX 5090.

The Download: Claude’s inner workings, and the future of world models

Model ReleasesDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. What Anthropic’s latest AI discovery does—and doesn’t—show —Ja

Together AI positions open-weight AI models as the enterprise moat for cost, control and IP

AgentsDGX agent

Enterprises racing to deploy AI at scale are discovering that the biggest constraint isn’t model capability anymore — it’s control. As agentic AI moves from experimentation into core business processe

13 Jul 2026

ICYMI - you can now build recursive language models (RLMs) with deepagents: the main agent can write custom code to recursively call subagen…

AgentsDGX agent

ICYMI - you can now build recursive language models (RLMs) with deepagents: the main agent can write custom code to recursively call subagents! this is super flexible -- the harness can take any shape

10 Jul 2026

AI Model Co-Design: Hardware-Friendly LLM Design

HardwareDGX agent

AI model co-design involves integrating model requirements into hardware planning to ensure memory hierarchies and networking stacks are purpose-built for specific systems . AI performance balances th

Answer Set Programming Energised! End-to-End Neurosymbolic Reasoning and Learning with ASP and Energy Based Models

Model ReleasesDGX agent

arXiv:2607.08136v1 Announce Type: new Abstract: We present a general neurosymbolic reasoning and learning methodology based on a modular integration of answer set programming with an energy based mode

Diagnosing Corruption-Induced Reliability Failures in Vision-Language Models

SafetyDGX agent

arXiv:2511.19032v2 Announce Type: replace Abstract: Visual corruptions can change vision--language model (VLM) behavior in ways that top-1 accuracy does not capture. A model may keep the same answer w

Do Egocentric Video-Language Models Capture Both Hand- and Object-Centric Cues?

ResearchDGX agent

arXiv:2607.08514v1 Announce Type: new Abstract: Hand-object interaction (HOI) recognition requires capturing both hand manipulations and object transformations. However, existing video-language models

End of an era: Decommissioning the original Model S & X assembly line in just 46 days

IndustryDGX agent

Tesla decommissioned the original Model S and Model X assembly line in 46 days, marking the end of an era for these vehicles that launched the company's expansion beyond the Roadster. This rapid trans

UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks

Model ReleasesDGX agent

arXiv:2607.08768v1 Announce Type: new Abstract: The rapid development of large language models and multimodal large language models has accelerated the emergence of proactive agents capable of operati

Unlocking Temporal Generalization in Hamiltonian Video Dynamics Models

ResearchDGX agent

arXiv:2607.07763v1 Announce Type: new Abstract: World models are typically trained to predict discrete-time physical dynamics with a fixed step size baked into the model weights, preventing prediction

9 Jul 2026

Amortized Inference for Correlated Discrete Choice Models via Equivariant Neural Networks

ResearchDGX agent

arXiv:2603.24705v3 Announce Type: replace-cross Abstract: Discrete choice models are fundamental tools in management science, economics, and marketing for understanding and predicting decision-making.

Comprehensive Evaluation of Large Language Model Responses: A Multi-Factor Scoring System

ResearchDGX agent

arXiv:2607.06940v1 Announce Type: cross Abstract: The remarkable performance of large language models (LLMs) in linguistic tasks underscores an urgent need for comprehensive evaluation of their respon

GemNav: Discrete-Token Visual Robot Navigation using a Multimodal Large Language Model

SafetyDGX agent

arXiv:2607.06882v1 Announce Type: cross Abstract: Visual navigation policies built on large pretrained models have so far followed a common recipe: a dedicated visual encoder, a bespoke action head, a

← Previous
1…126127128129130…1009
Next →