App alternatives on device transcription

Hello, I have been looking for apps that can transcribe my in person meetings using my phone as keeping laptop open might sometimes look weird. Ideally they also will clean up re organize the transcript if needed with local AI.

I need something that works offline as I don't want to risk data of confidential conversations leaking by sending it off to whatever provider.

I found this Notarius app recently but I find that while transcriptions are quite OK the AI part of it is a bit too simple... any ideas on alternatives that might work on iphone? Ideally works 100% in airplane mode.

reddit.com
u/Sharp-Translator6401 — 19 hours ago

App alternatives on device transcription

Hello, I have been looking for apps that can transcribe my in person meetings using my phone as keeping laptop open might sometimes look weird. Ideally they also will clean up re organize the transcript if needed with local AI.

I need something that works offline as I don't want to risk data of confidential conversations leaking by sending it off to whatever provider.

I found this Notarius app recently but I find that while transcriptions are quite OK the AI part of it is a bit too simple... any ideas on alternatives that might work on iphone? Ideally works 100% in airplane mode.

reddit.com
▲ 4 r/LocalAIServers+1 crossposts

Anyone with 4+ R9700? How do you combine them for inference?

I was recently playing with these GPUs to see how far I can push single stream inference on large models (need all cards to work) like DSV4 / Qwen 122B / ... but I am getting very different results depending on the model I use. Also I am getting very different card usage in terms of compute and power drained, for example for DSV4 cards are nearly idle and inference is slow.

I am not sure whether I am doing something wrong or the software is just not there for some architectures yet.

reddit.com
u/Sharp-Translator6401 — 4 days ago

Qwen 3.8 27b VS DSV4 Flash?

I know these are quite different model sizes but qwen is quite impressive and I am wondering if it is comparable to deepseek in terms of output quality... did anyone look into this?
I have purchased hardware to host deepseek at decent quant and speed but if qwen is already there... shall I give up on DS?

reddit.com
u/Sharp-Translator6401 — 5 days ago
▲ 49 r/LocalAIServers+2 crossposts

I did something crazy: connected 2x R9700 to Framework Desktop

I used the pcie 4x and one of the M.2 slots with respective raisers and external PSU.
Idea was to check: how far can I bring this platform, the 128GB vram is great but iGPU is kinda slow and bigger models struggle on it.
Also DSV4 flash sounds really nice but runs quite poorly on Strix Halo alone.

Here are some models I ran, as my workflows are mainly 'low concurrency' the tests were aimed at measuring single stream inference, not concurrent inference.

Please let me know if you would do anything different / try any other interesting model!

*Edit(i) The UD Q8 results of DS4 Flash (Q8_K_XL) are also computed now and its very close to Q4: pf 236 / tg 19.5

u/Sharp-Translator6401 — 6 days ago
▲ 6 r/LocalAIServers+1 crossposts

Anyone tried to train something on > 1 R9700?

Wanted to build a rig i can train small models on for some experiments and was thinking to get 3 or 4 R9700 as they are quite cost efficient but no idea how they perform for training

reddit.com
u/Sharp-Translator6401 — 14 days ago