Lilith Lilith.

From News

News · 2026-09-23

Remote GPUs help robots, but the network controls their reflexes

A Microsoft study found that moving inference from a robot to edge or cloud GPUs improves performance and battery life. It also exposes the hard constraint: tens of milliseconds of network delay already reduce the accuracy of physical work.

Read

News · 2026-09-21

Chimera framework advances drug discovery by ensembling diverse AI models

Microsoft Research and Novartis have introduced Chimera, a learning-to-rank system that shortens the retrosynthesis process in drug development.

Read

News · 2026-07-30

EvoLib replaces memory accumulation in LLMs with inference-time skill extraction

Microsoft Research introduced EvoLib, a framework that distills general skills from raw conversation history at inference time. This allows agents to improve runtime performance without requiring model retraining.

Read

News · 2026-07-30

Microsoft Echoverse builds evolving environments for agents, not static tests

Current agents fail in multi-step processes. Instead of static benchmarks, Microsoft Research trains them in simulations that dynamically evolve alongside task complexity.

Read

News · 2026-08-31

Microsoft’s GigaPath-Flash cuts the compute cost of pathology foundation models by 50x

Microsoft has distilled its massive pathology models into the Flash family. This delivers a 50x reduction in compute cost for whole-slide analysis, unlocking the possibility of repeated experiments across large patient populations.

Read

News · 2026-08-20

Skala 1.1 opens deep learning for quantum chemistry

Microsoft has released an update to its Skala model for quantum chemistry calculations. Version 1.1 makes more accurate molecular simulations accessible to a broader developer ecosystem and adds a living benchmark.

Read

News · 2026-08-12

Microsoft tests spatial reasoning: The new MindTopo benchmark

Microsoft introduced MindTopo, a benchmark for testing the spatial and topological reasoning of visual models. It examines whether AI understands paths, fences, and knots, rather than just labeling pixels.

Read

News · 2026-08-11

Microsoft adds precise measurement tools to Qwen for radiology

The Qwen3-VL-4B-Instruct model in Microsoft's research pipeline no longer just visually inspects X-rays, but triggers external tools to calculate dimensions. For threshold-dependent diagnostics, this means the end of merely guessing shapes.

Read

News · 2026-08-03

Microsoft opens Orchard: a greenhouse for training agents, not another orchestrator

Microsoft Research released Orchard, an open-source framework for scalable agentic model training. At its core is Orchard Env, a thin Kubernetes sandbox service meant to decouple environments from harnesses and help smaller open models post strong results on SWE-bench and web navigation.

Read

News · 2026-07-13

Microsoft moves Lean-checked crypto proofs into SymCrypt production code

Microsoft has published a SymCrypt branch with formal specifications and proofs for Rust implementations of SHA-3 and ML-KEM. The important move is not AI writing proofs, but proofs living next to production cryptographic code.

Read

News · 2026-06-29

Memora tackles agent memory by separating storage from retrieval

Microsoft Research describes Memora as a memory system for AI agents that currently need to reload or retrieve context repeatedly. The important move is separating what gets stored from how it is later retrieved.

Read

News · 2026-07-09

Aurora 1.5 pushes AI weather models toward energy and climate operations

Microsoft Research says Aurora 1.5 adds 22 more variables, hourly resolution and probabilistic ensemble forecasting. The interesting part is not prettier weather maps, but whether AI forecasts can support expensive operational decisions.

Read

News · 2026-07-08

Flint gives AI agents a shorter path from data to charts

Microsoft Research is showing Flint, an open-source visualization language that lets AI agents build charts from compact, human-editable specs. The interesting part is not another charting framework, but the layer between a prompt and the rendering library.

Read

News · 2026-07-02

Microsoft's Ire found LOTUSLITE where signatures failed

Microsoft showed Project Ire classifying a 253 KB LOTUSLITE DLL variant as malware even though VirusTotal showed only 1 of 72 vendors detecting it on May 28, 2026. The important part is not attribution, but an agent reading behavior instead of matching IOC lists.

Read

News · 2026-07-02

Microsoft turns AI explainability into a hypothesis engine for neuroscience

Microsoft Research described generative causal testing, a method that turns brain prediction models into readable hypotheses and checks them with fMRI. The interesting move is not prediction alone, but using AI to return testable theory to scientists.

Read

News · 2026-06-30

SkillOpt trains agent skills like text weights

Microsoft is showing SkillOpt, an optimizer that improves an agent skill file without changing model weights. For teams building agents, the important part is the validation gate, not another layer of prompt mysticism.

Read

News · 2026-06-24

Talos turns stored genomic data into a recurring shot at diagnosis

Microsoft Research and partners described Talos, an open-source tool for automated genomic reanalysis in rare disease. In a prospective cohort of almost 5,000 patients, it added diagnoses in 5.1 % of cases.

Read

News · 2026-05-21

MagenticLite combines small models, orchestration and local file access into one workflow without a frontier model

Microsoft Research describes MagenticLite, MagenticBrain and Fara1.5 as an agentic system optimized for small models that connects browser and local file system in a single workflow. The direction is practical: not one expensive model for everything, but orchestration of specialized components.

Read

News · 2026-05-28

Data Formulator 0.7 tries to rebuild enterprise data analytics around AI agents

Microsoft Research released Data Formulator 0.7, an analytics workspace where AI agents assist with exploration, transformation and visualization of enterprise data. The key question is whether the agent handles messy, permissioned data outside the demo.

Read