Aptura AI | Full-Time | MTS (Applied AI), MTS (SWE / Product) | London | ONSITE / HYBRID
We build the evaluation datasets and RL environments that make AI reliable where mistakes are expensive: finance, healthcare, and legal.
We design expert-curated training data, calibrated rubrics, and RL environments for nearly every frontier AI lab as well as startups pushing the frontier of what models can do.
We're a small London-based team running multiple active projects, so what you ship gets used immediately by labs, startups and internal domain experts. We recently launched SpreadsheetBench v2 (Kimi k3 model card) and work on ultra long-horizon tasks.
Aptura AI | Full-Time | MTS (Applied AI), MTS (SWE / Product) | London | ONSITE / HYBRID
We build the evaluation datasets and RL environments that make AI reliable where mistakes are expensive: finance, healthcare, and legal. We design expert-curated training data, calibrated rubrics, and RL environments for frontier AI labs and startups pushing the frontier of what models can do.
We're a small London-based team running multiple active projects, so what you ship gets used immediately by labs, startups and internal domain experts. We recently launched SpreadsheetBench v2 and work on ultra long-horizon tasks.
Not the author, but quickly ran a script over "Who wants to be hired?" from Jan-2021 to March-2023 and here is the resulting plot of the top-level comments (where indent="0"): https://www.arminbuilds.com/2021_01-2023_03_hired.png
So I was perhaps the inventor of "who wants to be hired" (I had asked to be able to post that thread, and was told that I needed to wait two days... anyway someone ran with it, and I am happy that you have made this thread.
For each size there is a ".en" and a multilingual one. I'm using the multilingual one -> you talk in German and receive a German transcript. On page 23 of the paper you can see the WER's of each language (https://cdn.openai.com/papers/whisper.pdf) I limited my app to the languages that had a max of around 20% WER.
Sure if the repo has a xcodeproj you can just open it in XCode, change the signing and improve on it. (Just make sure to always respect the licenses) If you want to play with whisper.cpp you can use this SwiftUI demo: https://github.com/ggerganov/whisper.cpp/tree/master/example...
Wow, thanks for the info and paper link! You are awesome. Also the assurance on how to fork projects. :) Good luck with your app! I'm also launching my Whisper related voice memo transcription app soon: https://apps.apple.com/app/wisprnote/id1671480366
We build the evaluation datasets and RL environments that make AI reliable where mistakes are expensive: finance, healthcare, and legal.
We design expert-curated training data, calibrated rubrics, and RL environments for nearly every frontier AI lab as well as startups pushing the frontier of what models can do.
We're a small London-based team running multiple active projects, so what you ship gets used immediately by labs, startups and internal domain experts. We recently launched SpreadsheetBench v2 (Kimi k3 model card) and work on ultra long-horizon tasks.
We're hiring:
* MTS, Applied AI — design the next benchmarks and RL environments — https://jobs.ashbyhq.com/aptura/c299e1f2-cdb4-4842-b5f3-8aa6...
* MTS, SWE / Product — build the platform behind every dataset — https://jobs.ashbyhq.com/aptura/cd6b6df2-1a75-442e-83b3-2ae0...