Dulus is the operating layer for inference 34 model providers unified into one runtime • 100+ inference backends through LiteLLM • 5. 9B tokens processed through Claude • 98. 8% prompt cache hit rate • Lookback compresses 2,000 conversation turns into a 20-turn inference window • Python Console a...

Source: [Hacker News](https://news.ycombinator.com/item?id=49136369)

Sponsored