Hacker Newsnew | past | comments | ask | show | jobs | submit | ArminRS's commentslogin

Aptura AI | Full-Time | MTS (Applied AI), MTS (SWE / Product) | London | ONSITE / HYBRID

We build the evaluation datasets and RL environments that make AI reliable where mistakes are expensive: finance, healthcare, and legal.

We design expert-curated training data, calibrated rubrics, and RL environments for nearly every frontier AI lab as well as startups pushing the frontier of what models can do.

We're a small London-based team running multiple active projects, so what you ship gets used immediately by labs, startups and internal domain experts. We recently launched SpreadsheetBench v2 (Kimi k3 model card) and work on ultra long-horizon tasks.

We're hiring:

* MTS, Applied AI — design the next benchmarks and RL environments — https://jobs.ashbyhq.com/aptura/c299e1f2-cdb4-4842-b5f3-8aa6...

* MTS, SWE / Product — build the platform behind every dataset — https://jobs.ashbyhq.com/aptura/cd6b6df2-1a75-442e-83b3-2ae0...


The swe/product position returns “The job you requested was not found”


Aptura AI | Full-Time | MTS (Applied AI), MTS (SWE / Product) | London | ONSITE / HYBRID

We build the evaluation datasets and RL environments that make AI reliable where mistakes are expensive: finance, healthcare, and legal. We design expert-curated training data, calibrated rubrics, and RL environments for frontier AI labs and startups pushing the frontier of what models can do.

We're a small London-based team running multiple active projects, so what you ship gets used immediately by labs, startups and internal domain experts. We recently launched SpreadsheetBench v2 and work on ultra long-horizon tasks.

We're hiring:

* MTS, Applied AI — design the next benchmarks and RL environments — https://jobs.ashbyhq.com/aptura/c299e1f2-cdb4-4842-b5f3-8aa6...

* MTS, SWE / Product — build the platform behind every dataset — https://jobs.ashbyhq.com/aptura/cd6b6df2-1a75-442e-83b3-2ae0...


Not the author, but quickly ran a script over "Who wants to be hired?" from Jan-2021 to March-2023 and here is the resulting plot of the top-level comments (where indent="0"): https://www.arminbuilds.com/2021_01-2023_03_hired.png


So I was perhaps the inventor of "who wants to be hired" (I had asked to be able to post that thread, and was told that I needed to wait two days... anyway someone ran with it, and I am happy that you have made this thread.

I propose the following:

WHO HAS BEEN HIRED and metrics on such.


Seems to anticorrelate pretty well


> It is priced at $0.002 per 1k tokens

So it would be $900 per day


For each size there is a ".en" and a multilingual one. I'm using the multilingual one -> you talk in German and receive a German transcript. On page 23 of the paper you can see the WER's of each language (https://cdn.openai.com/papers/whisper.pdf) I limited my app to the languages that had a max of around 20% WER.

Sure if the repo has a xcodeproj you can just open it in XCode, change the signing and improve on it. (Just make sure to always respect the licenses) If you want to play with whisper.cpp you can use this SwiftUI demo: https://github.com/ggerganov/whisper.cpp/tree/master/example...


Wow, thanks for the info and paper link! You are awesome. Also the assurance on how to fork projects. :) Good luck with your app! I'm also launching my Whisper related voice memo transcription app soon: https://apps.apple.com/app/wisprnote/id1671480366


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: