Unstructured
Open-source document intelligence: parsing, OCR, layout inference, table extraction, API reliability, and SDK work across 7 repos.
Medici Land Governance
Model-routing document pipelines for land records: OCR, VLMs, schema enforcement, and cost/accuracy escalation (Vertex AI, Gemini, Qwen3-VL, LoRA).
Sapient Logic
Cyber-defense NLP: semantic mapping of threat descriptions to MITRE ATT&CK, with embeddings, ranking, and mission OCR for sensitive environments.
PlusOne
Speech analytics at call-center scale: ASR, keyphrase extraction, sentiment, BERT modeling, and BigQuery ETL on thousands of calls.
DQLabs / Intellectyx
Semantic column-type detection for enterprise tabular data: Sherlock/Sato-style models that classify fields from data content, not headers.
SpeechLab
Event-driven serverless gateway for ASR, machine translation, and TTS on AWS, with cost and performance optimization.
EMCA / Energy Log Server
Time-series forecasting over Elasticsearch operational metrics: LSTM, ARIMA, and anomaly/prediction workflows.
Specific Diagnostics
Medical sensor-image analysis: converting raw imagery into colorimetric time-series with quality-control detection (TensorFlow, OpenCV).
Vasoactive Image Analysis
Microvascular before/after analysis: image registration, vessel skeletonization, diameter-change measurement, and PDF reporting.
AI Equity Research Analyst
Streamlit LLM app that reads 10-K filings via LlamaIndex sub-question retrieval and assembles structured equity-research sections.
Agentic Slides / Duarte AI
Agentic slide-generation and research automation prototypes with guardrails-oriented workflows, producing PPTX, text, and image artifacts.
Understanding Patient Conversation
Rasa conversational AI for medical questions: intent/entity recognition, chief-complaint matching, and dialogue management.
AI Customer Assistance
OCR/NLP system for building-defect communication: extracting and classifying specification documents into assistance workflows.
SHM Foundation
NLP email classification for nonprofit operations, with text-dataset preparation and requirements-driven ML project structure.