Editor’s Pick

Turla’s STOCKSTAY Backdoor: A New .NET Spy Tool Targeting Ukraine and Europe

Google Threat Intelligence Group's deep analysis of Turla's STOCKSTAY backdoor reveals a multi-component .NET implant that's been actively developed since 2022, targeting Ukrainian military and European diplomatic entities. The report includes code overlaps with the KAZUAR toolkit, YARA rules, and detailed operational timelines that are invaluable for threat hunters tracking Russian cyber espionage.

Read MoreTurla’s STOCKSTAY Backdoor: A New .NET Spy Tool Targeting Ukraine and Europe

Databricks’ Agent Cloud: Why Open Source and LTAP Matter for AI

Databricks co-founders Matei Zaharia and Reynold Xin unpack Omnigent (an open-source meta-harness above coding and enterprise agents), LTAP (their database bet for live transactional data in column-oriented formats), and why agent security, spend controls, and a common API matter more than ever. The thesis: traditional software gets rewritten once the data is in the right place and agents sit on top.

Read MoreDatabricks’ Agent Cloud: Why Open Source and LTAP Matter for AI

OpenAI and Broadcom unveil LLM-optimized inference chip Jalapeño

OpenAI and Broadcom unveiled Jalapeño, a custom LLM inference chip designed from scratch with substantially better performance per watt than current accelerators. Taped out in nine months using OpenAI's own models to accelerate chip design, it will deploy at gigawatt scale starting in 2026 as part of a multi-generation platform.

Read MoreOpenAI and Broadcom unveil LLM-optimized inference chip Jalapeño

NVIDIA NeMo AutoModel: 3.7x Faster MoE Fine-Tuning with One Import Change

NVIDIA NeMo AutoModel delivers 3.4-3.7x higher training throughput and 29-32% less GPU memory for MoE fine-tuning through a single import line change. By adding Expert Parallelism as a dedicated dimension, DeepEP fused dispatch, and TransformerEngine kernels on top of Transformers v5, it makes 550B-scale full fine-tuning feasible and produces standard HF checkpoints for downstream deployment.

Read MoreNVIDIA NeMo AutoModel: 3.7x Faster MoE Fine-Tuning with One Import Change

Local Model Triage for OpenClaw: Real-Time, Cost-Free PR Classification

If you own a powerful local machine like the NVIDIA GB10, using models like gemma-4-26b-a4b or qwen3.6-35b-a3b in an agent harness can give you real-time, cost-free triage of open-source contributions. The article provides concrete numbers on precision, recall, and throughput tradeoffs, plus a practical hybrid architecture that uses a cheap cloud audit loop ($9/month) to catch misses.

Read MoreLocal Model Triage for OpenClaw: Real-Time, Cost-Free PR Classification

AI Security After Codex and Claude Code: New Vulnerabilities and Guardrails

This episode from Gray Swan's cofounders explains why AI agents create a new class of security vulnerabilities that traditional cybersecurity cannot solve. Essential for engineers deploying Codex, Claude Code, or similar agents — prompt injection, automated red-teaming, and the lethal trifecta are must-understand concepts.

Read MoreAI Security After Codex and Claude Code: New Vulnerabilities and Guardrails