<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:media="http://search.yahoo.com/mrss/" xmlns:atom="http://www.w3.org/2005/Atom"  xmlns:dc="http://purl.org/dc/elements/1.1/">
	<channel>
		<title>overview for Milan_dr on Reddit</title>
		<description>overview for Milan_dr on Reddit</description>
		<pubDate>Thu, 23 Jul 2026 15:02:30 +0000</pubDate>
		<link>https://www.reddit.com/user/Milan_dr/</link>
		<atom:link href="https://fetchrss.com/feed/1l7r0rDKe8Ul1wmuwQCd7F9A.rss" rel="self" type="application/rss+xml" />
		<generator>https://fetchrss.com</generator>
		<image>
			<link>https://www.reddit.com/user/Milan_dr/</link>
			<url>https://res.cloudinary.com/dh2eofcns/provider/reddit.png</url>
			<title>overview for Milan_dr on Reddit</title>
		</image>

		<item>
			<guid isPermaLink="false">t1_ozb97zc</guid>
			<title>/u/Milan_dr on NanoGPT added new Mimo 2.5 and 2.5 Pro.</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v4gtwt/nanogpt_added_new_mimo_25_and_25_pro/ozb97zc/</link>
			<description><![CDATA[<div><p>Yep - so to be clear the reason we&#39;re not putting Crof as the provider for the &quot;general&quot; version of this model and doing this odd workaround is that we&#39;ve had a few too many reports of people that think the Crof versions of model are (slightly) worse in some ways, so we figure people can decide for themselves which to use.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Thu, 23 Jul 2026 16:36:23 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oz50x4k</guid>
			<title>/u/Milan_dr on The quality of NanoGPT’s models?</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v3j9og/the_quality_of_nanogpts_models/oz50x4k/</link>
			<description><![CDATA[<div><p>Thanks. Mimo 2.5 Pro we&#39;re aware - but frankly just unable to solve. Every provider we try for it has the same censoring issues (Xiaomi, Atlascloud, Novita, GMICloud) or quality issues (Deepinfra, Crof). If you&#39;ve tried these providers (or others) and are not getting the censoring would love to hear, because we&#39;re at our wits end on that one.</p> <p>In terms of GLM 5.2 - kind of same question, but any provider there where you <em>do</em> feel like it&#39;s good?</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Wed, 22 Jul 2026 19:27:53 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oz3saaz</guid>
			<title>/u/Milan_dr on The quality of NanoGPT’s models?</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v3j9og/the_quality_of_nanogpts_models/oz3saaz/</link>
			<description><![CDATA[<div><blockquote> <p>I’ve tested on both peak US/CN hours and off-peaks hours along with using providers that Nano may be routed to, the discounted ones for GLM 5.1/5.2 specifically. Is anyone else getting the same quality issue?</p> </blockquote> <p>Can I ask - which ones DO seem good/okay? So right now for example GLM 5.2 is a lot of Fireworks FP8, Z AI (supposedly FP8) and Novita (FP8).</p> <p>Someone a few days ago was claiming the ones that worked better were ones like Streamlake, Baidu etc, which we do not use (because of logging policies). </p> <p>Not saying you&#39;re wrong by the way, just trying to figure out what we can do here. Openrouter and us use largely the same providers (though that depends on what you have set in terms of logging policies), so trying to think what we can do.</p> <p>Same question for others in this thread actually - if people have been experimenting with PAYG/specific providers, would love to hear opinions on which ones DO work well for certain models. Because the most we can generally go off is &quot;what quantization do they run models on&quot; and &quot;what are their logging policies&quot;, but that&#39;s clearly incomplete.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Wed, 22 Jul 2026 16:24:17 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oyzp80y</guid>
			<title>/u/Milan_dr on New model: poolside/Laguna-S-2.1</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v2x36d/new_model_poolsidelagunas21/oyzp80y/</link>
			<description><![CDATA[<div><p>So what we see is that it has adaptive thinking - it does not always think, even when we try to &quot;force&quot; thinking.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Wed, 22 Jul 2026 01:33:08 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oyv8d9p</guid>
			<title>/u/Milan_dr on What happened to nano-gpt characters?</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v2hvv5/what_happened_to_nanogpt_characters/oyv8d9p/</link>
			<description><![CDATA[<div><p>So it&#39;s a work in progress - characters never was really a great integration so we kind of started from the ground up.</p> <p>Pushed this literally just yesterday, but obviously work in progress:</p> <p><a href="https://nano-gpt.com/roleplay">https://nano-gpt.com/roleplay</a></p> <p>We&#39;re considering making a roleplay.nanogpt.com, which we want to be a sort of.. simple version of SillyTavern, but then we also want to have all the extra options in there so that people can turn on an &quot;advanced&quot; version.</p> <p>Idea is to make it as easy as possible to get started, because while most here on the SillyTavern subreddit I think are quite deep into how all of this works and the many different options, we figure there are also lots of people that are just looking for a &quot;click character, get started&quot;. And we kind of have all the infrastructure at the ready for it.</p> <p>Either way - work in progress (I know I keep repeating that), hopefully more soon!</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Tue, 21 Jul 2026 13:14:03 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oyux2d5</guid>
			<title>/u/Milan_dr on Kimi k3 might actually be great</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v1ttyn/kimi_k3_might_actually_be_great/oyux2d5/</link>
			<description><![CDATA[<div><p>As far as we know it has mandatory reasoning.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Tue, 21 Jul 2026 12:14:52 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oyr08aa</guid>
			<title>/u/Milan_dr on How sillytavern handles context between chats and character names.</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v19732/how_sillytavern_handles_context_between_chats_and/oyr08aa/</link>
			<description><![CDATA[<div><p>In our case we don&#39;t do it (or not on purpose at least) - but it&#39;s technically possible that a provider in a way caches for.. similar words, maybe? </p> <p>Each chat is completely separate in the sense that we do not send a user identifier along, they&#39;re all just &quot;NanoGPT&quot;.</p> <p>We do try to maximize cache hits <em>within</em> a chat, so essentially we compare the hash of the first message in a chat (and system message) with other recent hashes, if it matches then it uses the same provider as was used for that message.</p> <p>So that could cause this for within-conversation, but across conversations that hash would be different hence no specific routing.</p> <p>Hope that makes sense, if not let me know.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Mon, 20 Jul 2026 21:08:43 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oymbfe0</guid>
			<title>/u/Milan_dr on I wish there was flat-rate subscription for a specific amount of tokens. Like a tier subscription</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v16rvj/i_wish_there_was_flatrate_subscription_for_a/oymbfe0/</link>
			<description><![CDATA[<div><p>For what it&#39;s worth you can also do that with us, just not in the subscription :)</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Mon, 20 Jul 2026 05:37:13 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oye39ct</guid>
			<title>/u/Milan_dr on Kimi K3's reasoning is insane.</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v04l51/kimi_k3s_reasoning_is_insane/oye39ct/</link>
			<description><![CDATA[<div><p>Even <a href="http://www.nanogpt.com">www.nanogpt.com</a> works nowadays - we need to move to that domain fully hah. Looks much better than with the dash. For now it&#39;s just a redirect.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Sun, 19 Jul 2026 00:33:57 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oybtsnd</guid>
			<title>/u/Milan_dr on ByteDance built an LLM specifically for roleplay #1 on OfoxAI's roleplay leaderboard.</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1uzyjeh/bytedance_built_an_llm_specifically_for_roleplay/oybtsnd/</link>
			<description><![CDATA[<div><p>Adding it in now.</p> <p>Edit: added</p> <p>Edit2: Will add it to subscription for now as well, should be live in ~10 minutes or so</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Sat, 18 Jul 2026 17:49:12 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oybsc6b</guid>
			<title>/u/Milan_dr on Gemma 4 31B Base or Gemma 4 31B IT, Which one is better for RP?</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1v00ylm/gemma_4_31b_base_or_gemma_4_31b_it_which_one_is/oybsc6b/</link>
			<description><![CDATA[<div><p>Far as I know there aren&#39;t providers for base, unless I&#39;ve missed them.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Sat, 18 Jul 2026 17:42:32 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oy40dqa</guid>
			<title>/u/Milan_dr on New open model - Inkling by Thinking Machines, available on NVDIA NIM</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1uy49v2/new_open_model_inkling_by_thinking_machines/oy40dqa/</link>
			<description><![CDATA[<div><p>Yup - was an issue on our side. They output it slightly differently, we didn&#39;t parse it correctly. Thanks for the poke on it.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Fri, 17 Jul 2026 15:28:49 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oy03ayr</guid>
			<title>/u/Milan_dr on Non-preview deepseek is out on the api?</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1uyiucq/nonpreview_deepseek_is_out_on_the_api/oy03ayr/</link>
			<description><![CDATA[<div><p>Were you using the deepseek v4 pro cheaper? Because the &quot;regular&quot; Deepseek v4 pro does not actually route via deepseek themselves, so it&#39;d have to be the open source providers somehow having early access.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Fri, 17 Jul 2026 00:48:39 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oxwht1t</guid>
			<title>/u/Milan_dr on New open model - Inkling by Thinking Machines, available on NVDIA NIM</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1uy49v2/new_open_model_inkling_by_thinking_machines/oxwht1t/</link>
			<description><![CDATA[<div><p>Yup! It is.</p> <p>We might add it to subscription - mostly depends on whether more providers start hosting it. So far it&#39;s slim pickings, to be honest with you.</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Thu, 16 Jul 2026 15:09:21 +0000</pubDate>
		</item>
		<item>
			<guid isPermaLink="false">t1_oxsxsc0</guid>
			<title>/u/Milan_dr on GLM 5.2 extremely slow?</title>
			<link>https://www.reddit.com/r/SillyTavernAI/comments/1uxaojs/glm_52_extremely_slow/oxsxsc0/</link>
			<description><![CDATA[<div><p>Hi.- when you say responses take extremely long to come through, do you mean it&#39;s a minute BEFORE the first response? As in, are you streaming and it takes that long to start returning? Or is this non stream?</p> </div>]]></description>
			<dc:creator>/u/Milan_dr</dc:creator>
			<pubDate>Thu, 16 Jul 2026 01:32:34 +0000</pubDate>
		</item>

	</channel>
</rss>