search
LLM models
Trends
- 1
NVIDIA/Model-Optimizer is an open-source Python library on GitHub that collects state-of-the-art model optimization techniques, including quantization, distillation, pruning, neural architecture search and speculative decoding. It compresses deep learning models so they run efficiently in deployment frameworks such as TensorRT-LLM, TensorRT and vLLM, improving inference speed. It is trending on GitHub's rankings with modest engagement, and the posts shown only describe the project itself, so there is no evidence of a specific event driving attention.
- 2
Pirate Face is a site or service being discussed on Hacker News under a headline claiming it 'rescues LLM models from deletion'. The post suggests it offers a way to preserve or recover large language models that might otherwise be removed, but the linked evidence gives no detail about how it works or who is behind it. The platform labels it as medicine, which does not obviously match the topic, and commenters' reactions are not visible, so sentiment and specifics are unclear from the posts alone.
Repos
- NandhaKishorM/laya Non-autoregressive System 1 decision engine. Typed choice, score and yes/no decisions over any text in a single forward
- unreallabsai/unreal-agent Async-first agent harness
- nokia-applied-research/AnyJev Turn any LLM into a Jev-style decision model: typed decisions, real probabilities, no training. (continue updating, welc
- browser-use/jev-ultrafast Fastest and cheapest web agent
- incoai/splash A local inference engine for Apple silicon, built around the model.
- mobile-next/mobile-mcp Model Context Protocol Server for Mobile Automation and Scraping (iOS, Android, Emulators, Simulators and Real Devices)
- NVIDIA/Model-Optimizer A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture se