<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Why is Splunk indexing our data in the wrong character encode? in Splunk Dev</title>
    <link>https://community.splunk.com/t5/Splunk-Dev/Why-is-Splunk-indexing-our-data-in-the-wrong-character-encode/m-p/421488#M7393</link>
    <description>&lt;P&gt;Splunk is indexing events in wrong format.&lt;/P&gt;

&lt;P&gt;On Splunk forwarder, I am seeing these errors:&lt;/P&gt;

&lt;PRE&gt;&lt;CODE&gt;WARN  UTF8Processor - Using charset UTF-8, as the monitor is believed over the raw text which may be UTF-16LE - data_source="C:\Program Files\SplunkUniversalForwarder\var\log\XXX.log", data_host="xxx", data_sourcetype="config"
&lt;/CODE&gt;&lt;/PRE&gt;

&lt;P&gt;A few events are indexed in the below format:&lt;/P&gt;

&lt;PRE&gt;&lt;CODE&gt;\xFF\xFEC\x00:\x00\\x00P\x00r\x00o
&lt;/CODE&gt;&lt;/PRE&gt;

&lt;P&gt;The input file data is in proper format which is output of Splunk btool cmd copied to file and ingested to Splunk.&lt;/P&gt;

&lt;P&gt;May I know how can we handle this?&lt;/P&gt;</description>
    <pubDate>Tue, 22 Jan 2019 21:42:41 GMT</pubDate>
    <dc:creator>ankithreddy777</dc:creator>
    <dc:date>2019-01-22T21:42:41Z</dc:date>
    <item>
      <title>Why is Splunk indexing our data in the wrong character encode?</title>
      <link>https://community.splunk.com/t5/Splunk-Dev/Why-is-Splunk-indexing-our-data-in-the-wrong-character-encode/m-p/421488#M7393</link>
      <description>&lt;P&gt;Splunk is indexing events in wrong format.&lt;/P&gt;

&lt;P&gt;On Splunk forwarder, I am seeing these errors:&lt;/P&gt;

&lt;PRE&gt;&lt;CODE&gt;WARN  UTF8Processor - Using charset UTF-8, as the monitor is believed over the raw text which may be UTF-16LE - data_source="C:\Program Files\SplunkUniversalForwarder\var\log\XXX.log", data_host="xxx", data_sourcetype="config"
&lt;/CODE&gt;&lt;/PRE&gt;

&lt;P&gt;A few events are indexed in the below format:&lt;/P&gt;

&lt;PRE&gt;&lt;CODE&gt;\xFF\xFEC\x00:\x00\\x00P\x00r\x00o
&lt;/CODE&gt;&lt;/PRE&gt;

&lt;P&gt;The input file data is in proper format which is output of Splunk btool cmd copied to file and ingested to Splunk.&lt;/P&gt;

&lt;P&gt;May I know how can we handle this?&lt;/P&gt;</description>
      <pubDate>Tue, 22 Jan 2019 21:42:41 GMT</pubDate>
      <guid>https://community.splunk.com/t5/Splunk-Dev/Why-is-Splunk-indexing-our-data-in-the-wrong-character-encode/m-p/421488#M7393</guid>
      <dc:creator>ankithreddy777</dc:creator>
      <dc:date>2019-01-22T21:42:41Z</dc:date>
    </item>
    <item>
      <title>Re: Why is Splunk indexing our data in the wrong character encode?</title>
      <link>https://community.splunk.com/t5/Splunk-Dev/Why-is-Splunk-indexing-our-data-in-the-wrong-character-encode/m-p/421489#M7394</link>
      <description>&lt;P&gt;HI,&lt;/P&gt;

&lt;P&gt;did you try to set the charset for your sourcetype?&lt;/P&gt;

&lt;P&gt;Usually if you change the CHARSET option in props.conf this will be fixed.&lt;BR /&gt;
Also be aware that the CHARSET option must be set on the UF or at input level - see more here &lt;A href="http://wiki.splunk.com/Where_do_I_configure_my_Splunk_settings"&gt;http://wiki.splunk.com/Where_do_I_configure_my_Splunk_settings&lt;/A&gt;&lt;/P&gt;

&lt;P&gt;Could be that you have to set it on indexer and UF, not sure about that, just try (&lt;A href="https://answers.splunk.com/answers/106700/seing-null-x00-bytes-in-indexed-data-from-log-file-in-windows.html"&gt;https://answers.splunk.com/answers/106700/seing-null-x00-bytes-in-indexed-data-from-log-file-in-windows.html&lt;/A&gt;)&lt;/P&gt;

&lt;P&gt;Would be someting like :&lt;/P&gt;

&lt;PRE&gt;&lt;CODE&gt;[&amp;lt;sourcetype&amp;gt;]
CHARSET = UTF16-LE
&lt;/CODE&gt;&lt;/PRE&gt;</description>
      <pubDate>Wed, 23 Jan 2019 09:16:20 GMT</pubDate>
      <guid>https://community.splunk.com/t5/Splunk-Dev/Why-is-Splunk-indexing-our-data-in-the-wrong-character-encode/m-p/421489#M7394</guid>
      <dc:creator>dkeck</dc:creator>
      <dc:date>2019-01-23T09:16:20Z</dc:date>
    </item>
    <item>
      <title>Re: Why is Splunk indexing our data in the wrong character encode?</title>
      <link>https://community.splunk.com/t5/Splunk-Dev/Why-is-Splunk-indexing-our-data-in-the-wrong-character-encode/m-p/530471#M7395</link>
      <description>&lt;DIV&gt;Hi Splunkers,&lt;BR /&gt;I have logs like&lt;/DIV&gt;&lt;DIV&gt;&lt;BR /&gt;&amp;lt;Header&amp;gt;&lt;BR /&gt;&amp;lt;Product&amp;gt;Microsoft SQL Server Reporting Services Version 2011.0110.6615.02 ((SQL11_SP3_QFE-CU).180109-2116 )&amp;lt;/Product&amp;gt;&lt;BR /&gt;&amp;lt;Locale&amp;gt;English ()&amp;lt;/Locale&amp;gt;&lt;BR /&gt;&amp;lt;TimeZone&amp;gt;Central Daylight Time&amp;lt;/TimeZone&amp;gt;&lt;BR /&gt;&amp;lt;Path&amp;gt;D:\Program Files\Microsoft SQL Server\MSRS11.CTSSRS2012\Reporting Services\Logfiles\ReportServerService__11_05_2020_14_52_11.log&amp;lt;/Path&amp;gt;&lt;BR /&gt;&amp;lt;SystemName&amp;gt;Avotrix69901&amp;lt;/SystemName&amp;gt;&lt;BR /&gt;&amp;lt;OSName&amp;gt;Microsoft Windows NT 6.2.9200&amp;lt;/OSName&amp;gt;&lt;BR /&gt;&amp;lt;OSVersion&amp;gt;6.2.9200&amp;lt;/OSVersion&amp;gt;&lt;BR /&gt;&amp;lt;ProcessID&amp;gt;3296&amp;lt;/ProcessID&amp;gt;&lt;BR /&gt;&amp;lt;Virtualization&amp;gt;Hypervisor&amp;lt;/Virtualization&amp;gt;&lt;BR /&gt;&amp;lt;/Header&amp;gt;&lt;BR /&gt;&amp;lt;ProcessorArchitecture&amp;gt;AMD64&amp;lt;/ProcessorArchitecture&amp;gt;&lt;BR /&gt;&amp;lt;ApplicationArchitecture&amp;gt;AMD64&amp;lt;/ApplicationArchitecture&amp;gt;&lt;BR /&gt;processing!ReportServer_0-51!1ed8!11/05/2020-14:52:11:: v VERBOSE: Mapping data reader successfully initialized.&lt;BR /&gt;library!ReportServer_0-51!2bc8!11/05/2020-14:52:11:: v VERBOSE: Transaction commit.&lt;BR /&gt;processing!ReportServer_0-51!1ed8!11/05/2020-14:52:11:: e ERROR: Throwing Microsoft.ReportingServices.ReportProcessing.ReportProcessingException: , Microsoft.ReportingServices.ReportProcessing.ReportProcessingException: There is no data for the field at position 3.;&lt;BR /&gt;runningjobs!ReportServer_0-51!2bc8!11/05/2020-14:52:11:: v VERBOSE: Thread pool settings: Available worker: 399, Max worker: 400, Available IO: 400, Max IO: 400&lt;BR /&gt;runningjobs!ReportServer_0-51!2bc8!11/05/2020-14:52:11:: v VERBOSE: Spawning new thread for a work item.&lt;BR /&gt;runningjobs!ReportServer_0-51!2bc8!11/05/2020-14:52:11:: v VERBOSE: ThreadJobContext.EndCancelableState&lt;BR /&gt;runningjobs!ReportServer_0-51!2bc8!11/05/2020-14:52:11:: v VERBOSE: ThreadJobContext.WaitForCancelException entered&lt;BR /&gt;runningjobs!ReportServer_0-51!2bc8!11/05/2020-14:52:11:: v&lt;/DIV&gt;&lt;DIV&gt;&lt;DIV&gt;&amp;nbsp;&lt;/DIV&gt;&lt;DIV&gt;And after indexing i am getting events like&lt;BR /&gt;\x00c\x00h\x00u\x00n\x00k\x00s\x00!\x00R\x00e\x00p\x00o\x00r\x00t\x00S\x00e\x00r\x00v\x00e\x005\x001\x00!\x002\x001\x00d\x000\x00!\x001\x001\x00/\x000\x005\x00/\x002\x000\x002\x000\x00-\x001\x004\x00:\x005\x002\x00:\x001\x002\x00:\x00:\x00 \x00v\x00 \x00V\x00E\x00R\x00B\x00O\x00S\x00E\x00:\x00 \x00R\x00e\x00t\x00r\x00i\x00e\x00v\x00e\x00d\x00 \x00s\x00e\x00g\x00m\x00e\x00n\x00t\x00 \x004\x003\x00f\x00b\x000\x009\x009\x00d\x00-\x00c\x006\x006\x004\x00-\x00e\x00a\x001\x001\x00-\x008\x001\x002\x00d\x00-\x000\x000\x002\x001\x005\x00a\x009\x00b\x000\x008\x00a\x00c\x00 \x00f\x00o\x00r\x00 \x00c\x00h\x00u\x00n\x00k\x00 \x004\x002\x00f\x00b\x000\x009\x009\x00d\x00-\x00c\x006\x006\x004\x00-\x00e\x00a\x001\x001\x00-\x008\x001\x002\x00d\x00-\x000\x000\x002\x001\x005\x00a\x009\x00b\x000\x008\x00a\x00c\x00 \x00f\x00r\x00o\x00m\x00 \x00t\x00h\x00e\x00 \x00s\x00e\x00g\x00m\x00e\x00n\x00t\x00&amp;nbsp;&lt;BR /&gt;&lt;BR /&gt;&lt;BR /&gt;&lt;BR /&gt;I had solved this issue using the below settings in props.conf&lt;BR /&gt;&lt;BR /&gt;&lt;BR /&gt;[MyOwnSourceType]&lt;BR /&gt;CHARSET = UTF16-LE&lt;/DIV&gt;&lt;/DIV&gt;</description>
      <pubDate>Mon, 23 Nov 2020 17:45:03 GMT</pubDate>
      <guid>https://community.splunk.com/t5/Splunk-Dev/Why-is-Splunk-indexing-our-data-in-the-wrong-character-encode/m-p/530471#M7395</guid>
      <dc:creator>VSIRIS</dc:creator>
      <dc:date>2020-11-23T17:45:03Z</dc:date>
    </item>
  </channel>
</rss>

