| | Even (very) noisy LLM evaluators are useful for improving AI agents (tensorzero.com) |
| 35 points by GabrielBianconi 3 months ago | past | 10 comments |
|
| | Stop comparing price per million tokens: the hidden LLM API costs (tensorzero.com) |
| 3 points by vrm 4 months ago | past | 2 comments |
|
| | We're building an automated AI engineer, and it works (tensorzero.com) |
| 3 points by GabrielBianconi 5 months ago | past |
|
| | Bandits in Your LLM Gateway (tensorzero.com) |
| 3 points by GabrielBianconi 10 months ago | past |
|
| | Is OpenAI's Reinforcement Fine-Tuning (RFT) Worth It? (tensorzero.com) |
| 4 points by GabrielBianconi 11 months ago | past |
|
| | 'Distealed' LLMs: smarter, 5-30x cheaper inference (tensorzero.com) |
| 3 points by anndvision on Sept 2, 2025 | past |
|
| | We raised $7.3M to build an open-source stack for industrial-grade LLM apps (tensorzero.com) |
| 1 point by GabrielBianconi on Aug 19, 2025 | past |
|
| | Fine-tuned small LLMs can beat large ones with programmatic data curation (tensorzero.com) |
| 53 points by GabrielBianconi on Aug 4, 2025 | past | 11 comments |
|
| | Automatically Evaluating AI Coding Assistants with Each Git Commit (tensorzero.com) |
| 3 points by vrm on July 5, 2025 | past |
|
| | Reverse Engineering Cursor's LLM Client (tensorzero.com) |
| 159 points by paulwarren on June 7, 2025 | past | 35 comments |
|
| | From NER to Agents: Does Automated Prompt Engineering Scale to Complex Tasks? (tensorzero.com) |
| 1 point by GabrielBianconi on April 9, 2025 | past |
|
| | Case Study: Automating Code Changelogs at a Large Bank with LLMs (tensorzero.com) |
| 1 point by GabrielBianconi on April 5, 2025 | past |
|
| | Think of LLM Applications as POMDPs – Not Agents (tensorzero.com) |
| 2 points by vrm on Feb 6, 2025 | past |
|