Add model and project credits to README
This commit is contained in:
26
README.md
26
README.md
@@ -49,9 +49,11 @@ you want to erase the evidence and start fresh.
|
||||
- **Inference:** llama.cpp compiled from source (pinned tag `b10333`) through the
|
||||
NDK. Pulled automatically at build time via CMake `FetchContent` — no submodule
|
||||
to init.
|
||||
- **Model:** `Qwen3.5-2B-Uncensored-HauhauCS-Aggressive-Q6_K.gguf` (~1.5 GB) is
|
||||
**downloaded on first launch** from Hugging Face into the app's private storage.
|
||||
It is *not* bundled in the APK. (Swap the quant in `ModelDownloader.kt`.)
|
||||
- **Model:**
|
||||
[`Qwen3.5-2B-Uncensored-HauhauCS-Aggressive`](https://huggingface.co/HauhauCS/Qwen3.5-2B-Uncensored-HauhauCS-Aggressive)
|
||||
in Q6_K format (~1.5 GB) is **downloaded on first launch** from Hugging Face
|
||||
into the app's private storage. It is *not* bundled in the APK. (Swap the
|
||||
quant in `ModelDownloader.kt`.)
|
||||
- **Persona:** set via `SYSTEM_PROMPT` in `MainActivity.kt`; sampling is a little
|
||||
hot (temp 0.9) for playful answers. The model is a Qwen3 "thinking" model, so the
|
||||
prompt disables reasoning (empty `<think></think>` prefill + `/no_think`) to keep
|
||||
@@ -99,6 +101,24 @@ manager, and tap it. First launch downloads the ~1.5 GB model over Wi-Fi.
|
||||
| Model URL / filename | `ModelDownloader.kt` |
|
||||
| Generation params (ctx, temp, tokens) | `MainActivity.kt` + `llama-jni.cpp` |
|
||||
|
||||
## Credits
|
||||
|
||||
tsjetpiti is small because it stands on the shoulders of some decidedly
|
||||
not-small projects:
|
||||
|
||||
- [Qwen3.5-2B](https://huggingface.co/Qwen/Qwen3.5-2B) by the Qwen team is the
|
||||
upstream base model.
|
||||
- [Qwen3.5-2B-Uncensored-HauhauCS-Aggressive](https://huggingface.co/HauhauCS/Qwen3.5-2B-Uncensored-HauhauCS-Aggressive)
|
||||
by [HauhauCS](https://huggingface.co/HauhauCS) is the model variant and GGUF
|
||||
quantization used by the app.
|
||||
- [llama.cpp](https://github.com/ggml-org/llama.cpp) by the ggml-org community
|
||||
provides the on-device inference engine and GGUF runtime.
|
||||
- [Android](https://developer.android.com/) and
|
||||
[Kotlin](https://kotlinlang.org/) provide the application platform and native
|
||||
app layer.
|
||||
- [Gradle](https://gradle.org/) and [CMake](https://cmake.org/) power the Kotlin
|
||||
and C++ build.
|
||||
|
||||
## TODO / notes
|
||||
|
||||
- **Branding:** launcher icon (adaptive + legacy densities) is generated from
|
||||
|
||||
Reference in New Issue
Block a user