Getting Data In

How to deal with data that was uploaded twice?

sudarshan391
Path Finder

Hi Experts,

I am now in a strange situation, we have a index in which we uploaded .csv files for every month and for previous month data has been uploaded two times. now splunk is showing duplicate entries.

Can someone please suggest how can I get through this situation?
I want to remove duplicate entries for last month from index.

Thanks.

Regards,
Sud

somesoni2
Revered Legend

You can use delete command to remove those duplicate records from any future search (it actually makes those records unsearchable but doesn't actually deletes it/removes from disk). See this for more information on the same.
https://docs.splunk.com/Documentation/Splunk/7.0.0/Indexer/RemovedatafromSplunk#Delete_events_from_s...

Please ensure that you run the search without delete command first and validate that you got only the records that you want to delete.

0 Karma

prosenjit2707
Explorer
0 Karma
Got questions? Get answers!

Join the Splunk Community Slack to learn, troubleshoot, and make connections with fellow Splunk practitioners in real time!

Meet up IRL or virtually!

Join Splunk User Groups to connect and learn in-person by region or remotely by topic or industry.

Get Updates on the Splunk Community!

Vibe-coding, AI, and Splunkcraft: Highlights from the .conf26 Builder Bar

If you stopped by the Builder Bar at .conf26, thank you! This year, we brought ...

Thanks for the Memories: .conf26 Took Learning to New Heights

Thank you, Splunk Community, for making .conf26 in Denver one for the books. From packed Splunk University ...

Best Practices: Splunk auto adjust pipeline queue

When you enable autoAdjustQueue in Splunk, maxSize should be understood as the queue size Splunk starts with ...