All Apps and Add-ons

website input charset problem

akdake
Explorer

I have create an input for scraping web-pages for data through Website Inputs add-on, and it work well.

However, the search result is some unreadable codes instead of Chinese character, , such Âé×핽ʽ£ºÈ«Â,
the html charset is gb2312, as following,

Also, the the CHARSET in props.conf is also gb2312

I have also tried charset HZ,utf-8,AUTO, in vain,

who can tell me why ? thanks .

Tags (2)
0 Karma

LukeMurphey
Champion

Please let me know if version 0.8 fixes your problem (or accept the answer so I know it worked).

0 Karma

LukeMurphey
Champion

This is a bug. The input isn't correctly determining the encoding of the page. I have a bug report created for and will get it fixed very soon.

Update:
This should work now as of version 0.8.

LukeMurphey
Champion

Is the site you are trying to get information from public? If so, could you share the selector you are using and the URL you are trying to load data from so that I could reproduce the issue?

akdake
Explorer

Yes, no search result was caputured after installing version, 0.8, I have tried this version on different Splunk demo,

0 Karma

LukeMurphey
Champion

Are you saying that it is no longer logging the results in Splunk?

0 Karma

akdake
Explorer

Many thanks, i update the add on to 0.8 , but it doesn't work, which cannot get any search result., pls confirm that.

0 Karma
Get Updates on the Splunk Community!

What's new in Splunk Cloud Platform 9.1.2312?

Hi Splunky people! We are excited to share the newest updates in Splunk Cloud Platform 9.1.2312! Analysts can ...

What’s New in Splunk Security Essentials 3.8.0?

Splunk Security Essentials (SSE) is an app that can amplify the power of your existing Splunk Cloud Platform, ...

Let’s Get You Certified – Vegas-Style at .conf24

Are you ready to level up your Splunk game? Then, let’s get you certified live at .conf24 – our annual user ...