Mimo V2.5-Pro open sourced
MiMo-V2.5-Pro is a fully open-sourced Mixture-of-Experts language model with 1.02T total parameters and 42B active parameters , available under the MIT License for commercial use, training, and fine-t
Knowledge catalogue
MiMo-V2.5-Pro is a fully open-sourced Mixture-of-Experts language model with 1.02T total parameters and 42B active parameters , available under the MIT License for commercial use, training, and fine-t
arXiv:2604.23733v1 Announce Type: new Abstract: Asking inquisitive questions while reading, and looking for their answers, is an important part in human discourse comprehension, curiosity, and creativ
No matter what kind of company you are...start making your internal company data legible to AI. Today. As a founder, you are essentially building two versions of your company: the one humans work in a
Now Available on Qdrant Cloud: GPU Indexing, Multi-AZ, and Audit Logging We’re excited to announce some Qdrand Cloud upgrades to address AI workloads that write continuously, must meet higher uptime S
@NVIDIA Nemotron 3 Nano Omni is now on Together AI. Enterprise multimodal AI — video, audio, image, documents & text — optimized for speed and scale. ✅ ~3B active params, 9x higher throughput ✅ Fully
arXiv:2412.04468v3 Announce Type: replace Abstract: Visual language models (VLMs) have made significant advances in accuracy in recent years. However, their efficiency has received much less attention
arXiv:2604.22837v1 Announce Type: cross Abstract: SAM-based dense trackers provide strong short-term mask propagation but remain fragile under long occlusion, fast motion, viewpoint change, and distra
arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th
arXiv:2510.01409v2 Announce Type: replace Abstract: System logs represent a valuable source of Cyber Threat Intelligence (CTI), capturing attacker behaviors, exploited vulnerabilities, and traces of m
Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where you don't need frontier intelligence. you want cheap, fast, an
arXiv:2604.24338v1 Announce Type: new Abstract: This paper evaluates an advanced jet trainer's utilization of artificial intelligence (AI)-based aircraft aerobatic maneuvers with the intention of deve
arXiv:2512.22113v2 Announce Type: replace-cross Abstract: Unresolved production cloud incidents cost an average of over $2M per hour. This paper introduces PRAXIS, an orchestrator that manages and dep
arXiv:2604.22868v1 Announce Type: cross Abstract: Visual planning represents a crucial facet of human intelligence, especially in tasks that require complex spatial reasoning and navigation. Yet, in m
arXiv:2604.23238v1 Announce Type: cross Abstract: Frontier models push the boundaries of what is learnable at extreme computational costs, yet distillation via sampling reasoning traces exposes closed
Open-source vector database startup Qdrant Solutions GmbH today announced three new enterprise-grade capabilities on its cloud service to address performance, availability and compliance requirements
arXiv:2603.25562v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) is increasingly used in LLM post-training because it can leverage a teacher model to provide dense supervision on
Unified identity security company Silverfort Inc. today announced that it has acquired Fabrix Security Ltd., an artificial intelligence-native identity security company, for an undisclosed price. Foun
arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to
arXiv:2604.24155v1 Announce Type: cross Abstract: The quest to align machine behavior with human values raises fundamental questions about the moral frameworks that should govern AI decision-making. M
arXiv:2604.22767v1 Announce Type: cross Abstract: Ethical discourse on AI in healthcare has focused predominantly on back-end concerns such as bias, fairness and explainability, while the front-end in
arXiv:2604.22789v1 Announce Type: cross Abstract: Organizations deploying AI-enabled Intelligent Transportation Systems face fragmented governance: ISO/IEC 42001 demands a certifiable management syste
arXiv:2604.22893v1 Announce Type: cross Abstract: Traditional data valuation methods based on ``row-count imes quality coefficient'' paradigms fail to capture the nuanced, nonlinear contributions that
Want to see which frontier models do the best on document understanding? Check out our ParseBench leaderboard on @kaggle! https://www.kaggle.com/benchmarks/llamaindex-org/parsebench For more details o
arXiv:2604.23724v1 Announce Type: cross Abstract: Expressway video anomaly detection is essential for safety management. However, identifying anomalies across diverse scenes remains challenging, parti
“Always keep a separate backup.” For sure, every experienced programmer should know this. 𝘉𝘶𝘵 𝘸𝘦 𝘫𝘶𝘴𝘵 𝘤𝘳𝘦𝘢𝘵𝘦𝘥 𝘢 𝘸𝘩𝘰𝘭𝘦 𝘨𝘦𝘯𝘦𝘳𝘢𝘵𝘪𝘰𝘯 𝘰𝘧 𝘷𝘪𝘣𝘦 𝘤𝘰𝘥𝘦𝘳𝘴 𝘸𝘩𝘰 𝘥𝘰𝘯’𝘵. And we now have a legion of synthetic coding
The Information: Analysis: as of late 2025, 79 of 500 tracked software companies including HubSpot, Adobe, and Salesforce adopted usage-based AI fees, more than doubling on 2024 — Dozens of enterprise
Sam Goldfarb / Wall Street Journal: Analysis of 100 actively traded software company loans since January 20 finds sector-wide price pressure as investors seek defensive moats against AI disruption — M
Google DeepMind announced a partnership with South Korea's Ministry of Science and ICT (MSIT) , marking the 10th anniversary of the historic AlphaGo match. Google will establish an AI Campus in Seoul
arXiv:2510.10254v2 Announce Type: replace Abstract: Recent advances in large generative models have shown that simple autoregressive formulations, when scaled appropriately, can exhibit strong zero-sh
Amazon Quick Flows enables users to build automations for repetitive or routine tasks using simple, everyday language prompts , with no technical expertise required. The service uses an AI-powered UI
arXiv:2604.22433v1 Announce Type: new Abstract: Heat exposure connects the built environment and public health, directly shaping the livability and sustainability of urban areas. Understanding the spa
arXiv:2601.19963v2 Announce Type: replace-cross Abstract: Training a high-performing neural decoder can be difficult when only limited data are available from a recording session. To address this chal
going to buy 2 rtx 6000 just because of the capabilities of local models becoming great! no more outsourcing of research to the api Xiaomi MiMo-V2.5 is now officially open-sourced! MIT License, suppor
I'm HIRING 5 more people to help me build the future of AI media. The Rundown is closing in on 3M active readers, and the demand is far outpacing what we can handle. 2,000 referral bonus if you help u
Loan processors spend 40–60% of their time reconciling income across tax returns, pay stubs, W-2s, and bank statements. We built an end-to-end pipeline that automates it with LlamaParse + the Claude A
microsoft/VibeVoice VibeVoice is Microsoft's Whisper-style audio model for speech-to-text, MIT licensed and with speaker diarization built into the model. Microsoft released it on January 21st, 2026 b
arXiv:2604.22239v1 Announce Type: cross Abstract: This paper introduces the task of analytical question answering over large, semi-structured document collections. We present MuDABench, a benchmark fo
arXiv:2604.22230v1 Announce Type: cross Abstract: Benchmark hacking refers to tuning a machine learning model to score highly on certain evaluation criteria without improving true generalization or fa
People keep blaming the users for the growing number of vibe-coded reasoners. Doing so is *half* right. Users *are*screwing up - by letting vibe coded stuff access their files, without proper backups,
arXiv:2511.14427v3 Announce Type: replace-cross Abstract: Effective contact-rich manipulation requires robots to synergistically leverage vision, force, and proprioception. However, Reinforcement Lear
.@swyx on whether AI infrastructure has finally stabilized: Media Always enjoy getting to chat with @swyx on our annual cross-episode with @latentspacepod on the state of AI. We hit on what’s shifted,
arXiv:2510.07632v2 Announce Type: replace Abstract: Frontier AI models have achieved remarkable progress, yet recent studies suggest they struggle with compositional reasoning, often performing at or
The artificial intelligence arms race has spent the last two years obsessed with a duopoly of constraints: the desperate hunt for Nvidia Corp. silicon and the grueling wait for grid-scale megawatts. I
arXiv:2604.22260v1 Announce Type: cross Abstract: Urban transportation systems face growing safety challenges that require scalable intelligence for emerging smart mobility infrastructures. While rece
Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate' and have that thing be infinitely scalable. See below image
I am increasingly bullish on open source harnesses (like OpenCode) not because they will be better than SOTA closed harnesses, but because they will never pull shady stuff like what Claude Code and ot
Opened a llama.cpp discussion about whether custom GBNF grammars can compose with tool calls in llama-server. Right now tools work alone, grammar works alone, but tools+grammar doesn't. If you use lla
The community can now download pre-quantized weights from MLX community repo on HF thanks to @LambdaAPI Model collection: https://huggingface.co/collections/mlx-community/deepseek-v4 DeepSeek-V4-Flash
GPT-5.5 prompting guide Now that GPT-5.5 is available in the API, OpenAI have released a wealth of useful tips on how best to prompt the new model. Here's a neat trick they recommend for applications
Launching pyptx — a Python DSL for writing NVIDIA PTX kernels. One PTX instruction = one Python call. Write pure PTX in Python. Direct Hopper + Blackwell support: wgmma, TMA, tcgen05, mbarriers. JAX +
Local models do seem likely to create an explosion of new use cases. Local compute >> cloud compute. This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚♀️ Qwen3.6 27B runn
Fireworks AI posted a question to their audience asking what features or capabilities they would like to see developed next, with a specific mention of full parameter tuning as a potential option. Thi
4⃣4⃣4⃣4⃣ Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit
arXiv:2604.21688v1 Announce Type: cross Abstract: The IC3 algorithm represents the state-of-the-art (SOTA) hardware model checking technique, owing to its robust performance and scalability. A signifi
This article describes a technique for accelerating reinforcement learning (RL) rollouts using distribution-aware speculative decoding, which can achieve up to 50% speedup improvements. The method lik
arXiv:2604.20932v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are increasingly deployed in sensitive domains such as healthcare and law, where they rely on private, do
arXiv:2504.03476v2 Announce Type: replace Abstract: Accurate lumbar spine segmentation is crucial for diagnosing spinal disorders. Existing methods typically use coarse-grained segmentation strategies
Developers can build with DeepSeek V4 through NVIDIA GPU-accelerated endpoints on build.nvidia.com, with hosted endpoints providing a fast way to prototype before moving to self-hosted deployment. Dee
arXiv:2604.21352v1 Announce Type: new Abstract: Mental health challenges are increasing worldwide, straining emotional support services and leading to counselor overload. This can result in delayed re
Cloud-hosted models will become more niche/seldom-used as much cheaper local models fulfill that which 90% of consumers need from AI. Lots of models to buy/use, but very few hardware options upon whic