RankLLMs Changelog

Real-time timeline of AI model benchmarks, LLM price drops, new features, and free credit deals.

AI Models

Published: Claude 3.7 Sonnet vs DeepSeek R1 vs GPT-4.5: Master 2026 AI Benchmark & Agent Guide

Comprehensive 2026 AI model comparison evaluating Claude 3.7 Sonnet, DeepSeek R1, and GPT-4.5 on SWE-bench Lite, MATH-500, pricing per 1M tokens, and agent execution.

Feature

Secure Serverless DeepSeek V3 Summarizer & LRU Cache Implemented

Launched an ultra-fast serverless Cloudflare API proxy for DeepSeek V3 summarization featuring a 50-capacity LRU Cache Data Structure to protect API tokens.

Free Credits

Published: Free AI API Credits & Developer Tiers (2026 Verified List)

Claim active free AI API credits, free GPU compute tiers, and promo codes for DeepSeek V3, Claude 3.5, OpenAI, Cloudflare Workers AI, and OpenRouter.

Deals

Launched Free AI API Credits & Limited-Time Deals Hub

Introduced a dedicated section and SEO hub (/free-credits) for verified free LLM API credits, developer GPU tier promos, and limited-time deals.

MDX

Published: Top LLM Benchmark Leaderboard 2026: Compare Speed, ELO & Pricing

Interactive LLM benchmark comparison guide. Analyze speed, cost per 1M tokens, ELO ratings, and performance metrics across top AI models.

News

Published: Freebuff AI Review: Is It Safe & Legit? (Free Cursor Alternative)

In-depth Freebuff AI Coding Agent review. Is Freebuff safe and legit? Tested on security, privacy, free tier limits, Claude 3.5 Sonnet proxy speed, and VS Code setup.

Claude

Published: Claude Code vs Gemini CLI vs Copilot CLI: Best Terminal AI Agent (2026)

Benchmark comparison of Claude Code, Gemini CLI, and GitHub Copilot CLI. Tested on multi-file refactoring, token costs, context speed, and terminal integration.

Claude

Published: Claude 4 Sonnet vs GPT 5.1 Benchmark: Speed, Cost & Coding Test (2026)

Direct benchmark comparison: Claude 4 Sonnet vs GPT 5.1. Compare SWE-bench verified scores, MATH-500 accuracy, token pricing per million, and context windows.

Gemini

Published: Gemini 3 Pro vs Gemini 2.5 Pro: Benchmark, Latency & 2M Context Test

Comparing Google's Gemini 3 Pro vs Gemini 2.5 Pro. Breakdown of TTFT latency (~290ms), 2 million token context stability, video processing, and code execution.

Grok

Published: Grok 4 vs Claude 4 vs Gemini 2.5 Pro: 2026 AI Model Comparison

Comprehensive AI comparison of Grok 4, Claude 4, and Gemini 2.5 Pro. Compare real-time web search, coding capability, multimodal performance, and context scale.