|
Chunking strategy for governments and internal org docs
|
|
2
|
59
|
August 5, 2026
|
|
Technical question: provider-attested served revision for Inference Providers
|
|
1
|
51
|
August 2, 2026
|
|
Voice agent latency degrades after turn 7-8 despite fixed system prompt + limited history — looking for mitigation ideas beyond what we've already tried
|
|
1
|
79
|
July 28, 2026
|
|
Real-time voice agents with local LLMs: the latency problem nobody fully solves
|
|
4
|
173
|
July 22, 2026
|
|
I am developing a non-neural conversational AI model
|
|
6
|
171
|
July 17, 2026
|
|
DPO Training ruins my model’s conversational coherence
|
|
3
|
364
|
July 15, 2026
|
|
Building a Centralized Vector Database Pipeline for Multi-Project RAG
|
|
0
|
62
|
July 10, 2026
|
|
Distinguish between thinking and responding during generation
|
|
1
|
81
|
July 10, 2026
|
|
Help with DeepSeek-V3-0324 Model Download
|
|
6
|
545
|
July 9, 2026
|
|
Setup: Ollama serving llama3.1:8b-instruct-q4_K_M, chat completions API (/api/chat), num_ctx=4096, temperature 0.2-0.3
|
|
1
|
64
|
July 8, 2026
|
|
Wav2vec2 / WavLM audio classifier stuck at chance (33%) — only training the head
|
|
5
|
114
|
July 2, 2026
|
|
Huggingface/text-embeddings-inference, cpu bug
|
|
0
|
38
|
June 24, 2026
|
|
I built an open source VAD that beats Silero, Pyannote, and WebRTC on noisy audio with 93% accuracy — no GPU required
|
|
0
|
168
|
June 21, 2026
|
|
Custom semantic representation ("bryła") beats raw text in 24/27 configs — built solo on an RTX 2060, looking for feedback
|
|
6
|
228
|
June 14, 2026
|
|
🚀 New tool for AI manga creators: **MangaBuilder** (buildmanga.com)
|
|
1
|
497
|
June 14, 2026
|
|
AI Career Choice
|
|
3
|
151
|
June 12, 2026
|
|
What are the core components required to build a robust AI agent in 2026?
|
|
2
|
216
|
June 11, 2026
|
|
CUDA support added - Pre-generation knowledge-boundary estimator
|
|
1
|
108
|
June 9, 2026
|
|
Fine-Tuning an SLM for a Low-Resource Language
|
|
7
|
274
|
June 6, 2026
|
|
Pre-generation knowledge-boundary estimator
|
|
0
|
34
|
June 5, 2026
|
|
Agent Valve survey: where agents get blocked
|
|
0
|
15
|
June 3, 2026
|
|
Collaborators, and feedback on 1st development
|
|
2
|
94
|
June 2, 2026
|
|
Finetuning a Reasoning LLM with Supervised or Reinforcement Learning?
|
|
1
|
347
|
June 2, 2026
|
|
Physical Modelling of sim2real SO101 Arm Project
|
|
1
|
201
|
May 30, 2026
|
|
Training lora for LTX2.3 voice / sound only
|
|
4
|
769
|
May 27, 2026
|
|
Which framework is better for chatbot development: LangChain or LlamaIndex?
|
|
1
|
74
|
May 21, 2026
|
|
An error in docker: failed to create shim task
|
|
2
|
142
|
May 20, 2026
|
|
Issue while quantizing Gemma 4 E2B/E4B - TypeError: torch.finfo() requires a floating point input type. Use torch.iinfo to handle 'torch.finfo'
|
|
2
|
115
|
May 17, 2026
|
|
A use-case example for data transfer between LLM chat threads and maintaining architectural continuity across systems
|
|
0
|
29
|
May 15, 2026
|
|
How do I improve my Ai Vtuber?
|
|
1
|
104
|
May 15, 2026
|