Joaquin Garcia Personal
Personal projects by Joaquin Garcia
About Joaquin Garcia Personal
Personal projects by Joaquin Garcia
Projects (15)
MiniMax H3
freemiumMiniMax H3 is an AI video generator that produces 4-15 second 2K clips with native 32 kHz stereo audio generated in the same pass, from $0.13 per second.
Audio Convert
freemiumAudio Convert turns audio and video into editable text with OpenAI Whisper. Upload, record, or paste a URL, then export TXT, SRT, DOCX, or JSON.
GPT Transcribe
freemiumGPT Transcribe is an online AI speech to text workspace that turns audio and video into searchable, timestamped transcripts with speaker labels, running on OpenAI Whisper across 100+ languages. Export as TXT, SRT, VTT, JSON, PDF or DOCX.
Krea 2 Turbo
freemiumKrea 2 Turbo is an AI image generator running the eight-step distilled Krea 2 model — 1024px drafts in about two seconds, plus image to image, LoRA sliders and 2K export.
Speech Notes
freemiumSpeech Notes is a browser AI speech to text workspace that turns recordings, files and live audio into editable, searchable transcripts in 100+ languages.
Image To Image
freemiumImage To Image is an AI photo editor that transforms an uploaded photo with image-to-image style transfer, background swap, and plain-English edits in 4K.
Get Img AI
freemiumGet Img AI is a free AI image generator and AI photo editor: text-to-image, image-to-image, and 9-reference fusion in 4K across multiple top AI models.
Speech Text
freemiumSpeech Text is an online AI speech to text workspace that transcribes audio, video, live recordings and media URLs into editable, exportable transcripts.
Muse Image
freemiumMuse Image is an AI image editor and AI Art Generator for text to image, image to image editing, and reference-guided visuals in one browser workspace.
Fish Voice
freemiumFish Voice is an AI voice generator that turns editable text into expressive speech, with a multilingual public voice library and consent-first voice cloning.
Omni Voice
freemiumOmni Voice is a free AI voice generator that turns scripts into natural text to speech, with consent-based voice cloning, voice design, and EN/ZH/JA/KO support.