my-pal-mcp-server

Author	SHA1	Message	Date
Fahad	23353734cd	Support for allowed model restrictions per provider Tool escalation added to `analyze` to a graceful switch over to codereview is made when absolutely necessary	2025-06-14 10:56:53 +04:00
Fahad	2c805d6637	Fixed mock comparison error	2025-06-14 09:34:56 +04:00
Fahad	746380eb7f	Renamed setup script to avoid confusion (https://github.com/BeehiveInnovations/zen-mcp-server/issues/35 ) Further fixes to tests Pass O3 simulation test when keys are not set, along with a notice Updated docs on testing, simulation tests / contributing Support for OpenAI o4-mini and o4-mini-high	2025-06-14 09:28:20 +04:00
Fahad	c5f682c7b0	Fix tests to work with effective auto mode changes - Added autouse fixture to mock provider availability in tests - Updated test expectations to match new auto mode behavior - Fixed mock provider capabilities to return proper values - Updated claude continuation tests to set default model - All 256 tests now passing 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-14 02:43:29 +04:00
Fahad	eb388ab2f2	Categorize tools into 'model capabilities categories' to help determine which type of model to pick when in auto mode Encourage Claude to pick the best model for the job automatically in auto mode Lots of new tests to ensure automatic model picking works reliably based on user preference or when a matching model is not found or ambiguous Improved error reporting when bogus model is requested and is not configured or available	2025-06-14 02:17:06 +04:00
Fahad	8ac5bbb5af	Fixed workspace path mapping Refactoring Improved system prompts, more generalized Home folder protection and detection Retry logic for gemini	2025-06-14 00:26:59 +04:00
Fahad	26b22a1d53	Simplified /workspace to map to a project scoped WORKSPACE_ROOT	2025-06-13 20:49:37 +04:00
Fahad	fb69ebebe4	Lint	2025-06-13 16:13:02 +04:00
Fahad	a7b27b285c	Cleanup and confirm tests pass	2025-06-13 16:09:40 +04:00
Fahad	048ebf90bf	Cleanup	2025-06-13 16:05:21 +04:00
Fahad	f44ca326ef	Breaking change: openrouter_models.json -> custom_models.json * Support for Custom URLs and custom models, including locally hosted models such as ollama * Support for native + openrouter + local models (i.e. dozens of models) means you can start delegating sub-tasks to particular models or work to local models such as localizations or other boring work etc. * Several tests added * precommit to also include untracked (new) files * Logfile auto rollover * Improved logging	2025-06-13 15:22:09 +04:00
Fahad	a641159a67	Use consistent terminology Remove test folder from .gitignore for live simulation test to pass	2025-06-13 09:28:33 +04:00
Fahad	b16f85979b	Use consistent terminology	2025-06-13 09:06:12 +04:00
Fahad	0e36fcbc69	Final cleanup	2025-06-13 07:12:29 +04:00
Fahad	2cdb92460b	WIP - OpenRouter model configuration registry - Model definition file for users to be able to control - Additional tests - Update instructions	2025-06-13 06:33:12 +04:00
Fahad	cd1105b741	WIP - OpenRouter model configuration registry - Model definition file for users to be able to control - Update instructions	2025-06-13 05:52:26 +04:00
Fahad	a19055b76a	WIP - OpenRouter model configuration registry - Model definition file for users to be able to control - Update instructions	2025-06-13 05:52:16 +04:00
Fahad	52b45f2b03	WIP - OpenRouter support and related refactoring	2025-06-12 22:17:11 +04:00
Fahad	22093bbf18	Fixed tests	2025-06-12 21:00:53 +04:00
Fahad	3aedb16101	Use the new Gemini 2.5 Flash Updated to support Thinking Tokens as a ratio of the max allowed Updated tests Updated README	2025-06-12 20:46:54 +04:00
Fahad	354a0fae0b	Fixed tests	2025-06-12 13:51:22 +04:00
Fahad	79af2654b9	Use the new flash model Updated tests	2025-06-12 13:44:09 +04:00
Fahad	7462599ddb	Simplified thread continuations Fixed and improved tests	2025-06-12 12:47:02 +04:00
Fahad	fb66825bf6	Rebranding, refactoring, renaming, cleanup, updated docs	2025-06-12 10:40:43 +04:00
Fahad	9a55ca8898	WIP lots of new tests and validation scenarios Simulation tests to confirm threading and history traversal Chain of communication and branching validation tests from live simulation Temperature enforcement per model	2025-06-12 09:35:05 +04:00
Fahad	2a067a7f4e	WIP major refactor and features	2025-06-12 07:14:59 +04:00
Fahad	22a3fb91ed	feat: Add comprehensive dynamic configuration system v3.3.0 ## Major Features Added ### 🎯 Dynamic Configuration System - Environment-aware model selection: DEFAULT_MODEL with 'pro'/'flash' shortcuts - Configurable thinking modes: DEFAULT_THINKING_MODE_THINKDEEP for extended reasoning - All tool schemas now dynamic: Show actual current defaults instead of hardcoded values - Enhanced setup workflow: Copy from .env.example with smart customization ### 🔧 Model & Thinking Configuration - Smart model resolution: Support both shortcuts ('pro', 'flash') and full model names - Thinking mode optimization: Only apply thinking budget to models that support it - Flash model compatibility: Works without thinking config, still beneficial via system prompts - Dynamic schema descriptions: Tool parameters show current environment values ### 🚀 Enhanced Developer Experience - Fail-fast Docker setup: GEMINI_API_KEY required upfront in docker-compose - Comprehensive startup logging: Shows current model and thinking mode defaults - Enhanced get_version tool: Reports all dynamic configuration values - Better .env documentation: Clear token consumption details and model options ### 🧪 Comprehensive Testing - Live model validation: New simulator test validates Pro vs Flash thinking behavior - Dynamic configuration tests: Verify environment variable overrides work correctly - Complete test coverage: All 139 unit tests pass, including new model config tests ### 📋 Configuration Files Updated - docker-compose.yml: Fail-fast API key validation, thinking mode support - setup-docker.sh: Copy from .env.example instead of manual creation - .env.example: Detailed documentation with token consumption per thinking mode - .gitignore: Added test-setup/ for cleanup ### 🛠 Technical Improvements - Removed setup.py: Fully Docker-based deployment (no longer needed) - REDIS_URL smart defaults: Auto-configured for Docker, still configurable for dev - All tools updated: Consistent dynamic model parameter descriptions - Enhanced error handling: Better model resolution and validation ## Breaking Changes - Removed setup.py (Docker-only deployment) - Model parameter descriptions now show actual defaults (dynamic) ## Migration Guide - Update .env files using new .env.example format - Use 'pro'/'flash' shortcuts or full model names - Set DEFAULT_THINKING_MODE_THINKDEEP for custom thinking depth 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-11 20:10:25 +04:00
Fahad	b0f17f741f	Fixed invalid test assumptions	2025-06-11 18:49:30 +04:00
Fahad	e8df6a7a31	Comments	2025-06-11 17:18:40 +04:00
Fahad	780000f9c9	Lots of tests with live simulation to validate conversation continuation / preservation work across requests	2025-06-11 17:16:05 +04:00
Fahad	c90ac7561e	Lots of tests with live simulation to validate conversation continuation / preservation work across requests	2025-06-11 17:03:09 +04:00
Fahad	ac763e0213	More tests	2025-06-11 14:34:51 +04:00
Fahad	98eab46abf	WIP - improvements to token usage tracking, simulator added for live testing, improvements to file loading	2025-06-11 13:24:59 +04:00
Fahad	5a94737516	Fix conversation history duplication and optimize file embedding This major refactoring addresses critical bugs in conversation history management and significantly improves token efficiency through intelligent file embedding: Key Improvements: • Fixed conversation history duplication bug by centralizing reconstruction in server.py • Added intelligent file filtering to prevent re-embedding files already in conversation history • Centralized file processing logic in BaseTool._prepare_file_content_for_prompt() • Enhanced log monitoring with better categorization and file embedding visibility • Updated comprehensive test suite to verify new architecture and edge cases Architecture Changes: • Removed duplicate conversation history reconstruction from tools/base.py • Conversation history now handled exclusively by server.py:reconstruct_thread_context • All tools now use centralized file processing with automatic deduplication • Improved token efficiency by embedding unique files only once per conversation Performance Benefits: • Reduced token usage through smart file filtering • Eliminated redundant file embeddings in continued conversations • Better observability with detailed debug logging for file operations 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-11 11:40:12 +04:00
Fahad	94f542c76a	fix: critical conversation history bug and improve Docker integration This commit addresses several critical issues and improvements: 🔧 Critical Fixes: - Fixed conversation history not being included when using continuation_id in AI-to-AI conversations - Fixed test mock targeting issues preventing proper conversation memory validation - Fixed Docker debug logging functionality with Gemini tools 🐛 Bug Fixes: - Docker compose configuration for proper container command execution - Test mock import targeting from utils.conversation_memory.* to tools.base.* - Version bump to 3.1.0 reflecting significant improvements 🚀 Improvements: - Enhanced Docker environment configuration with comprehensive logging setup - Added cross-tool continuation documentation and examples in README - Improved error handling and validation across all tools - Better logging configuration with LOG_LEVEL environment variable support - Enhanced conversation memory system documentation 🧪 Testing: - Added comprehensive conversation history bug fix tests - Added cross-tool continuation functionality tests - All 132 tests now pass with proper conversation history validation - Improved test coverage for AI-to-AI conversation threading ✨ Code Quality: - Applied black, isort, and ruff formatting across entire codebase - Enhanced inline documentation for conversation memory system - Cleaned up temporary files and improved repository hygiene - Better test descriptions and coverage for critical functionality 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-11 08:53:45 +04:00
Fahad	14ccbede43	Fixes https://github.com/BeehiveInnovations/gemini-mcp-server/issues/6	2025-06-11 07:09:28 +04:00
Fahad	ce23a996c6	Conversation threading test fixes	2025-06-10 20:53:20 +04:00
Fahad	f5060367a0	WIP - communication memory	2025-06-10 19:16:51 +04:00
Fahad	032e783efb	Groundings added	2025-06-10 14:39:14 +04:00
Fahad	ba8f7192c3	refactor: rename think_deeper to thinkdeep for brevity - Renamed `think_deeper` tool to `thinkdeep` for shorter, cleaner naming - Updated all imports from ThinkDeeperTool to ThinkDeepTool - Updated all references from THINK_DEEPER_PROMPT to THINKDEEP_PROMPT - Updated tool registration in server.py - Updated all test files to use new naming convention - Updated README documentation to reflect new tool names - All functionality remains the same, only naming has changed This completes the tool renaming refactor for improved clarity and consistency. 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-10 12:38:38 +04:00
Fahad	5f8ed3aae8	refactor: rename review tools for clarity and consistency - Renamed `review_code` tool to `codereview` for better naming convention - Renamed `review_changes` tool to `precommit` to better reflect its purpose - Updated all tool descriptions to remove "Triggers:" sections and improve clarity - Updated all imports and references throughout the codebase - Renamed test files to match new tool names - Updated server.py tool registrations - All existing functionality preserved with improved naming This refactoring improves code organization and makes tool purposes clearer. 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-10 12:30:06 +04:00
Fahad	67f18ef3c9	refactor: rename debug_issue tool to debug for brevity - Rename debug_issue.py to debug.py - Update tool name from 'debug_issue' to 'debug' throughout codebase - Update all references in server.py, tests, and README - Keep DebugIssueTool class name for backward compatibility - All tests pass with the renamed tool This makes the tool name shorter and more consistent with other tool names like 'chat' and 'analyze'. The functionality remains exactly the same. 🤖 Generated with Claude Code Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-10 11:43:47 +04:00
Fahad	21b0470aef	style: remove trailing whitespace in test_docker_path_integration.py 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-10 10:02:01 +04:00
Fahad	788d1fa9d3	fix: prevent double translation of already-translated Docker paths Added check in translate_path_for_environment() to detect and skip already-translated container paths (those starting with /workspace). This prevents the function from attempting to translate paths like: - /workspace/src/main.py -> /inaccessible/outside/mounted/volume/workspace/src/main.py Now it correctly handles: - Host path: /Users/.../src/main.py -> /workspace/src/main.py (translation) - Container path: /workspace/src/main.py -> /workspace/src/main.py (no change) Added comprehensive test to verify double-translation prevention. 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-10 09:58:25 +04:00
Fahad	27add4d05d	feat: Major refactoring and improvements v2.11.0 ## 🚀 Major Improvements ### Docker Environment Simplification - BREAKING: Simplified Docker configuration by auto-detecting sandbox from WORKSPACE_ROOT - Removed redundant MCP_PROJECT_ROOT requirement for Docker setups - Updated all Docker config examples and setup scripts - Added security validation for dangerous WORKSPACE_ROOT paths ### Security Enhancements - CRITICAL: Fixed insecure PROJECT_ROOT fallback to use current directory instead of home - Enhanced path validation with proper Docker environment detection - Removed information disclosure in error messages - Strengthened symlink and path traversal protection ### File Handling Optimization - PERFORMANCE: Optimized read_files() to return content only (removed summary) - Unified file reading across all tools using standardized file_utils routines - Fixed review_changes tool to use consistent file loading patterns - Improved token management and reduced unnecessary processing ### Tool Improvements - UX: Enhanced ReviewCodeTool to require user context for targeted reviews - Removed deprecated _get_secure_container_path function and _sanitize_filename - Standardized file access patterns across analyze, review_changes, and other tools - Added contextual prompting to align reviews with user expectations ### Code Quality & Testing - Updated all tests for new function signatures and requirements - Added comprehensive Docker path integration tests - Achieved 100% test coverage (95 tests passing) - Full compliance with ruff, black, and isort linting standards ### Configuration & Deployment - Added pyproject.toml for modern Python packaging - Streamlined Docker setup removing redundant environment variables - Updated setup scripts across all platforms (Windows, macOS, Linux) - Improved error handling and validation throughout ## 🔧 Technical Changes - Removed: `_get_secure_container_path()`, `_sanitize_filename()`, unused SANDBOX_MODE - Enhanced: Path translation, security validation, token management - Standardized: File reading patterns, error handling, Docker detection - Updated: All tool prompts for better context alignment ## 🛡️ Security Notes This release significantly improves the security posture by: - Eliminating broad filesystem access defaults - Adding validation for Docker environment variables - Removing information disclosure in error paths - Strengthening path traversal and symlink protections 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-10 09:50:05 +04:00
Fahad	7ea790ef88	fix: Docker path translation for review_changes and code deduplication - Fixed review_changes tool to properly translate host paths to container paths in Docker - Prevents "No such file or directory" errors when running in Docker containers - Added proper error handling with clear messages when paths are inaccessible refactor: Centralized token limit validation across all tools - Added _validate_token_limit method to BaseTool to eliminate code duplication - Reduced ~25 lines of duplicated code across 5 tools (analyze, chat, debug_issue, review_code, think_deeper) - Maintains exact same error messages and behavior feat: Enhanced large prompt handling - Added support for prompts >50K chars by requesting file-based input - Preserves MCP's ~25K token capacity for responses - All tools now check prompt size before processing test: Added comprehensive Docker path integration tests - Tests for path translation, security validation, and error handling - Tests for review_changes tool specifically with Docker paths - Fixed failing think_deeper test (updated default from "max" to "high") chore: Code quality improvements - Applied black formatting across all files - Fixed import sorting with isort - All tests passing (96 tests) - Standardized error handling follows MCP TextContent format The changes ensure consistent behavior across all environments while reducing code duplication and improving maintainability. 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-10 07:20:24 +04:00
Fahad	53303f86be	feat: enhance review_changes with dynamic file requests - Add instruction for Gemini to request files when needed - Add comprehensive tests for files parameter functionality - Test file request instruction presence/absence based on context - Run all tests, ruff, and black formatting Now review_changes can both accept context files and allow Gemini to request additional files during review for better validation. 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-09 21:43:45 +04:00
Fahad	783ba73181	refactor: cleanup and comprehensive documentation Major changes: - Add comprehensive documentation to all modules with detailed docstrings - Remove unused THINKING_MODEL config (use single GEMINI_MODEL with thinking_mode param) - Remove list_models functionality (simplified to single model configuration) - Rename DEFAULT_MODEL to GEMINI_MODEL for clarity - Remove unused python-dotenv dependency - Fix missing pydantic in setup.py dependencies Documentation improvements: - Document security measures in file_utils.py (path validation, sandboxing) - Add detailed comments to critical logic sections - Document tool creation process in BaseTool - Explain configuration values and their impact - Add comprehensive function-level documentation Code quality: - Apply black formatting to all files - Fix all ruff linting issues - Update tests to match refactored code - All 63 tests passing 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-09 19:04:24 +04:00
Fahad	fd6e2f9b64	refactor: rename review_pending_changes to review_changes - Renamed tool from review_pending_changes to review_changes for brevity - Enhanced tool descriptions for better MCP auto-discovery - Updated all references throughout codebase including: - Tool implementation (tools/review_changes.py) - Test files (tests/test_review_changes.py) - Server registration and imports - Documentation in README.md - Tool prompts in prompts/tool_prompts.py - Enhanced review_changes description to emphasize pre-commit usage - All tests pass, linting and formatting checks pass 🤖 Generated with [Claude Code](https://claude.ai/code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-06-09 14:37:03 +04:00
Fahad	dc366d3a23	refactor: remove unused TOOL_TRIGGERS dead code - Remove unused TOOL_TRIGGERS dictionary from config.py - Remove associated test_tool_triggers test case - TOOL_TRIGGERS was not used anywhere in the codebase - MCP automatically discovers tools through descriptions in list_tools handler - All tests pass (64 tests), ruff clean, black formatted	2025-06-09 14:24:59 +04:00

1 2

76 Commits