Antonio
AI Research Engineer specializing in the research, development, and deployment of solutions based on LLMs, SLMs, fine-tuning, RAG, and agentic AI architectures.
Recent posts
-
GLM-5.2: I thought Sonnet 4.5 and similar open models were enough, but…
Honestly, it had been a long time since a model surprised me like this — not since the end of 2025. No, I do not mean it surprised me because I liked a response or was impressed by a benchmark. No, and in this article you will understand where I am going with this. I remember the launch of Sonnet 4.5 as if it were yesterday. It was one of those moments when the developer community noticed a consid…
- claude
- claude-sonnet
- +8
-
Agentic AI, SLMs, and Why Models Above US$0.50 Output per 1M Tokens Are Equivalent to Burning Money
June 2026. If you are still paying more than fifty cents per million output tokens in any agentic pipeline, forgive my frankness: you are literally burning money. No, this is not an exaggeration on my part, but simply the relationship between API costs and absurd usage — because, indeed, LLMs should be used absurdly, whether you are a solo researcher like me or a startup. And here, the reason has …
- deepseek
- llm
- +4