Add model and project credits to README
This commit is contained in:
26
README.md
26
README.md
@@ -49,9 +49,11 @@ you want to erase the evidence and start fresh.
|
|||||||
- **Inference:** llama.cpp compiled from source (pinned tag `b10333`) through the
|
- **Inference:** llama.cpp compiled from source (pinned tag `b10333`) through the
|
||||||
NDK. Pulled automatically at build time via CMake `FetchContent` — no submodule
|
NDK. Pulled automatically at build time via CMake `FetchContent` — no submodule
|
||||||
to init.
|
to init.
|
||||||
- **Model:** `Qwen3.5-2B-Uncensored-HauhauCS-Aggressive-Q6_K.gguf` (~1.5 GB) is
|
- **Model:**
|
||||||
**downloaded on first launch** from Hugging Face into the app's private storage.
|
[`Qwen3.5-2B-Uncensored-HauhauCS-Aggressive`](https://huggingface.co/HauhauCS/Qwen3.5-2B-Uncensored-HauhauCS-Aggressive)
|
||||||
It is *not* bundled in the APK. (Swap the quant in `ModelDownloader.kt`.)
|
in Q6_K format (~1.5 GB) is **downloaded on first launch** from Hugging Face
|
||||||
|
into the app's private storage. It is *not* bundled in the APK. (Swap the
|
||||||
|
quant in `ModelDownloader.kt`.)
|
||||||
- **Persona:** set via `SYSTEM_PROMPT` in `MainActivity.kt`; sampling is a little
|
- **Persona:** set via `SYSTEM_PROMPT` in `MainActivity.kt`; sampling is a little
|
||||||
hot (temp 0.9) for playful answers. The model is a Qwen3 "thinking" model, so the
|
hot (temp 0.9) for playful answers. The model is a Qwen3 "thinking" model, so the
|
||||||
prompt disables reasoning (empty `<think></think>` prefill + `/no_think`) to keep
|
prompt disables reasoning (empty `<think></think>` prefill + `/no_think`) to keep
|
||||||
@@ -99,6 +101,24 @@ manager, and tap it. First launch downloads the ~1.5 GB model over Wi-Fi.
|
|||||||
| Model URL / filename | `ModelDownloader.kt` |
|
| Model URL / filename | `ModelDownloader.kt` |
|
||||||
| Generation params (ctx, temp, tokens) | `MainActivity.kt` + `llama-jni.cpp` |
|
| Generation params (ctx, temp, tokens) | `MainActivity.kt` + `llama-jni.cpp` |
|
||||||
|
|
||||||
|
## Credits
|
||||||
|
|
||||||
|
tsjetpiti is small because it stands on the shoulders of some decidedly
|
||||||
|
not-small projects:
|
||||||
|
|
||||||
|
- [Qwen3.5-2B](https://huggingface.co/Qwen/Qwen3.5-2B) by the Qwen team is the
|
||||||
|
upstream base model.
|
||||||
|
- [Qwen3.5-2B-Uncensored-HauhauCS-Aggressive](https://huggingface.co/HauhauCS/Qwen3.5-2B-Uncensored-HauhauCS-Aggressive)
|
||||||
|
by [HauhauCS](https://huggingface.co/HauhauCS) is the model variant and GGUF
|
||||||
|
quantization used by the app.
|
||||||
|
- [llama.cpp](https://github.com/ggml-org/llama.cpp) by the ggml-org community
|
||||||
|
provides the on-device inference engine and GGUF runtime.
|
||||||
|
- [Android](https://developer.android.com/) and
|
||||||
|
[Kotlin](https://kotlinlang.org/) provide the application platform and native
|
||||||
|
app layer.
|
||||||
|
- [Gradle](https://gradle.org/) and [CMake](https://cmake.org/) power the Kotlin
|
||||||
|
and C++ build.
|
||||||
|
|
||||||
## TODO / notes
|
## TODO / notes
|
||||||
|
|
||||||
- **Branding:** launcher icon (adaptive + legacy densities) is generated from
|
- **Branding:** launcher icon (adaptive + legacy densities) is generated from
|
||||||
|
|||||||
Reference in New Issue
Block a user