Infrastructure & renewable energy investment professional (20+ years, PMP) who builds private, local-first AI tools. Zürich, Switzerland · junqueira.ch · LinkedIn
I have spent two decades developing, financing and delivering hydropower, solar and gas assets across Latin America, Africa and Europe. Today I combine that domain knowledge with hands-on AI engineering.
Turning expert knowledge into private AI. Energy and infrastructure companies hold know-how and trade secrets they cannot paste into a public chatbot. I am building the pipeline to make that knowledge usable by AI on their own infrastructure:
recordings / PDFs / documents → clean Markdown corpus → verified revision (tcqa) → RAG (retrieval-augmented generation) → later: fine-tuning a model on the company's own material
All tools were built with Claude (Anthropic) as coding partner. I write the requirements, test on real data and iterate. Every earlier version is kept in each repo's archive/ folder.
| Project | What it does | Size |
|---|---|---|
| TranscriptLab | Whisper transcription, YouTube/podcast capture and Markdown polishing into a RAG-ready corpus | ~27k lines, 70+ versions |
| MD Converter | PDF/Office/HTML/EPUB to Markdown with 6+ local engines, isolated environments and smoke tests. Windows .exe download in Releases | ~5.5k lines, 49 versions |
| tcqa | Offline fidelity checker: proves a corrected speech-to-text transcript changed only what its log says, before the text enters a RAG or fine-tuning corpus | 14 checks, 203 tests, CI on Ubuntu and Windows |
| RadioSave | Scheduled recorder for online radio: captures interviews and author programs unattended, as raw material for the transcript pipeline | ~3.7k lines |
| SRT Translator | Structure-preserving subtitle translation with Claude, with quality control, live pricing, cost estimates and a one-screen desktop GUI. Windows .exe download in Releases | ~2.7k lines, 10 versions |
- Write a clear brief first, then iterate in small versioned steps
- Give the AI rules it must obey (module boundaries, naming, safety checklists); MediaClinic shows this written contract in practice
- Test on real data, log everything, fix the root cause
- Keep the history public
- Keep confidential data local and secrets out of the code
Not everything here is AI. I am a civil engineer by training, and one older project belongs on this page.
dlearn-ppd is my 2001 civil-engineering graduation project at UNESP Bauru, rebuilt for current machines. It comes from research on dynamic structural analysis: DLEARN, the finite-element program published with T.J.R. Hughes' textbook, extended with Prof. Heitor M. Bottura's Hermitian time-integration algorithms. The pre-processor I wrote in Turbo Pascal generates DLEARN's input files from a question-and-answer dialogue. In 2026 I rebuilt DLEARN with gfortran and, with Claude, added a tested Python rewrite of the pre-processor. Checking my 2001 program against the Fortran exposed three bugs in it, documented in the repo.
One of my hobbies is keeping my films and series in my own offline library: rips of the DVDs and Blu-rays I have bought, served from a home media server or NAS to Plex, Kodi, Emby or Jellyfin. I use the tools below every day, and they are built for people who run the same kind of library.
Why keep your own copies? Streaming catalogues change without asking you. Titles leave with little or no notice, sometimes in the middle of a series, and the subscription price does not go down when the catalogue shrinks. A disc on my shelf and a file on my disk are still there next year: in the quality I chose, with the audio and subtitle languages I want, with no account, no licence server and no internet connection needed.
The catch is that a big library only works if it is tidy. Missing artwork, wrong IDs, broken metadata files and inconsistent genres make Plex, Kodi, Emby or Jellyfin show the wrong movie, or none at all. These tools keep it healthy:
| Tool | What it does |
|---|---|
| MediaClinic | For Plex, Kodi, Emby and Jellyfin users: scans a movie library, shows every metadata, artwork and video problem in one colour-coded table and fixes the common ones safely. It checks Kodi .nfo and Emby/Jellyfin movie.xml files, so it also suits Plex with an NFO add-on. Windows .exe download in Releases |
| Kodi Files Generator | Builds Kodi NFO/XML files from a CSV and checks folder names |
Subtitles are part of the same job: I use SRT Translator (listed under AI projects above) to translate the subtitles of the films I own, and the same tool turns the subtitles of videos into text for my transcript corpus.
This is about keeping what I have bought, not about piracy. Rules on copying discs differ by country, so check yours.
Python · Tkinter · Whisper · MarkItDown · Docling · ffmpeg · yt-dlp · Claude API · Markdown pipelines · RAG concepts · Kodi/Plex NFO and XML metadata
Open to senior roles in infrastructure and renewables, and to conversations about private AI for energy companies. 📧 [email protected] · 🌐 junqueira.ch


