DataSci Ocean
Posts
Tags
Categories
English
繁體中文
English
Light
Dark
Auto
DataSci Ocean
Cancel
Posts
Tags
Categories
Light
Dark
Auto
English
繁體中文
English
Inference Optimization
2026
AgentOpt: How to Optimize Client-Side LLM Agents and Cut API Costs by 67%
04-17
Beyond HyDE: How ReDE-RF Makes RAG 10x Faster by "Judging" Instead of "Writing"
02-03
Stop Using Giant LLMs for Everything: Why NVIDIA Research Says Small Language Models (SLMs) Are the Future of AI Agents
01-19
2025
Stop AI Hallucinations Early: How Meta's DeepConf Uses Token Confidence to Cut Inference Costs by 80%
12-28
2024
Better & Faster Large Language Models via Multi-token Prediction
07-18