Portfolio
Systems I designed, built, and run.
queryForge™
A financial-data platform I built and operate. I use it for my own market analysis.
Project Overview
queryForge pulls live and historical market data through the OpenBB API and serves it as charts and a JSON API. It isn't a product I sell — it's a working system that shows how I build.
Key Features
- Real-time stock and crypto quotes
- Multiple chart types (line, bar, candlestick)
- JSON API responses for integrations
- Redis caching
- Responsive web interface
- RESTful API architecture
Technical Stack
Frontend
Django Templates, React Components, Modern CSS
Backend
Django, FastAPI, PostgreSQL, Redis
External APIs
OpenBB Financial Data API
DevOps
GitHub Actions (CI/CD), Postman (API Testing), Unit Tests
Deployment
DigitalOcean
queryLane™ Data Platform
Big data processing platform built on the Hadoop ecosystem
Project Description
queryLane is a data lake on Hadoop, Hive, and Spark for large-scale ingestion, processing, and storage.
Capabilities
- Multi-terabyte data ingestion
- Distributed processing with Spark
- SQL-like queries with Hive
- Schema evolution support
- Integration with data warehouses
- Data quality validation pipelines
Technical Implementation
Infrastructure
WSL Ubuntu Environment with Hadoop Cluster
Data Processing
- Apache Hadoop (HDFS)
- Apache Hive (Data Warehouse)
- Apache Spark (Processing)
- Python (ETL Scripts)
Use Cases
- Log file analysis
- Time-series data storage
- Data lake architecture
- Batch processing workflows
Monitoring Data Pipeline
Python-based data cleaning and transformation pipeline
Project Overview
A production-grade data cleaning pipeline designed to process years of monitoring data. This project demonstrates best practices in data quality, testing, and maintainability.
Key Features
- Automated data validation
- Missing value handling
- Outlier detection
- Data type standardization
- Incremental processing
- Error logging and reporting
Technologies
- Python 3.x
- Pandas for data manipulation
- Pytest for unit testing
- Great Expectations (validation)
- Logging and monitoring
- CI/CD integration
Quality Assurance
- Automated testing in CI/CD pipeline
- Data quality metrics and reporting
- Documentation and code comments
Additional Projects & Capabilities
Time Series Analysis
Multi-year monitoring data analysis with trend forecasting and anomaly detection.
Python • Pandas • Scikit-learn • Statsmodels
Power BI Dashboards
Interactive dashboards for business intelligence and operational monitoring.
Power BI • DAX • Data Modeling
API Development
RESTful APIs with FastAPI for data access and integrations.
FastAPI • PostgreSQL • Redis • OpenAPI
Data Science Projects
Statistical analysis, predictive modeling, and machine learning implementations.
Python • R • Jupyter • Scikit-learn
Want something like this?
The same approach, applied to your operations. No pitch, just perspective.