feat: Implement smart time-based rate limiting v0.1.8

🎯 Major Rate Limiting Improvement - Research-Friendly Design

New Rate Limiting System:
✅ 12 calls per 5-minute window (vs 8 per session)
✅ 8 calls per 30-second burst protection
✅ Automatic time-based reset (no LM Studio restarts needed)
✅ Clear feedback with remaining calls and reset times

Key Benefits:
- Supports extended research sessions without interruption
- Prevents LLM spam while allowing legitimate research workflows
- Users can work continuously without restarting LM Studio
- Intelligent burst protection prevents overwhelming websites

Technical Implementation:
- Sliding window algorithm with timestamp tracking
- Dual-layer protection: burst + window limits
- Automatic cleanup of expired call history
- User-friendly error messages with precise wait times

This addresses the core user feedback that session-based limits were too restrictive for normal research use cases while maintaining responsible web scraping practices.

Breaking Change: Rate limiting behavior changed from session-based to time-based
Migration: No action needed - new system is more permissive
This commit is contained in:
Jay
2025-08-10 23:47:42 -05:00
parent d380736ea0
commit e1f6fff3fe
4 changed files with 65 additions and 14 deletions
+20 -1
View File
@@ -1,4 +1,4 @@
# 🌐 WebFetch.MCP v0.1.7
# 🌐 WebFetch.MCP v0.1.8
**Live Web Access for Your Local AI — Tunable Search & Clean Content Extraction**
@@ -119,6 +119,25 @@ In LM Studio:
| DEBUG | false | Debug logging |
| DETAILED_LOG | true | Detailed log output |
# ⏱️ Smart Rate Limiting
WebFetch.MCP uses intelligent time-based rate limiting designed for real research workflows:
### **📊 Rate Limits:**
- **12 calls per 5-minute window** - Generous limit for research sessions
- **8 calls per 30-second burst** - Prevents LLM spam while allowing quick queries
- **Automatic reset** - No need to restart LM Studio between research sessions
### **🎯 Why This Works Better:**
- ✅ **Research-friendly** - Supports extended research sessions
- ✅ **Anti-spam protection** - Prevents runaway LLM tool calling
- ✅ **No restarts needed** - Limits reset automatically over time
- ✅ **Clear feedback** - Shows remaining calls and reset times
### **📈 Example Usage Patterns:**
- **Quick research**: 5-8 rapid calls, then brief pause
- **Extended research**: 12 calls spread over 5 minutes
- **Continuous work**: Limits reset as you work, no interruption
# 📊 Example Usage
**Search**
```