Getting Data In

Best way to index .csv files with some common fields

simonbuskens
Engager

Hi
I have a series of .csv files (1 for each month) where the first 100 fields are the same, but after that there are about 4 or 5 fields that are specific to that month only.

What is the best way to add that data into Splunk?

Will I be able to add them all to the same index, or do they need their own index (so then I have to search across them all)

Tags (3)
1 Solution

esix_splunk
Splunk Employee
Splunk Employee

For your sourcetype, define the fieldnames in the header in all of your csv's in props.conf.

Then, define additional fields in that header for the sourcetype. E.g.,

fields = alwaysherefield1, alwaysherefield2, alwaysherefield...30, sometimesherefield31, sometimesherefield32, sometimesherefield33

If those fields dont exists, there wont be a value set. In the CSVs that have the fields, the values will be set. Then in your search, you can rename the fields to whatever you want them to be.

Refer to here for props.conf examples :
http://docs.splunk.com/Documentation/Splunk/6.2.1/Data/Extractfieldsfromfileheadersatindextime

View solution in original post

esix_splunk
Splunk Employee
Splunk Employee

For your sourcetype, define the fieldnames in the header in all of your csv's in props.conf.

Then, define additional fields in that header for the sourcetype. E.g.,

fields = alwaysherefield1, alwaysherefield2, alwaysherefield...30, sometimesherefield31, sometimesherefield32, sometimesherefield33

If those fields dont exists, there wont be a value set. In the CSVs that have the fields, the values will be set. Then in your search, you can rename the fields to whatever you want them to be.

Refer to here for props.conf examples :
http://docs.splunk.com/Documentation/Splunk/6.2.1/Data/Extractfieldsfromfileheadersatindextime

asifhj
Path Finder

Files name are same or different? if the file name are different you can put in same index and can write queries according to source field.

0 Karma

kml_uvce
Builder

you can add in same index and source name will be different and you can search based on source name

kamal singh bisht
0 Karma

simonbuskens
Engager

won't the source name be the same (if they are saved in the same folder)

0 Karma

esix_splunk
Splunk Employee
Splunk Employee

Source name will be path+filename, sourcetype will be whatever you defined it to be.

0 Karma

asifhj
Path Finder

Nope, folders cannot have two files with same name.

0 Karma
Got questions? Get answers!

Join the Splunk Community Slack to learn, troubleshoot, and make connections with fellow Splunk practitioners in real time!

Meet up IRL or virtually!

Join Splunk User Groups to connect and learn in-person by region or remotely by topic or industry.

Get Updates on the Splunk Community!

Splunk Auto Ingestion Parallel Pipeline Scaling

Why this feature matters Many Splunk environments experience ingestion pressure long before the host is fully ...

Best Practices: Splunk auto adjust pipeline queue

When you enable autoAdjustQueue in Splunk, maxSize should be understood as the queue size Splunk starts with ...

Announcing Modern Navigation: A New Era of Splunk User Experience

We are excited to introduce the Modern Navigation feature in the Splunk Platform, available to both cloud and ...