Here is a simple solution: Replace all \n with \\n before saving to CSV. This will preserve the newline characters.
df.loc[:, "Column_Name"] = df["Column_Name"].apply(lambda x: x.replace('\n', '\\n'))
df.to_csv("df.csv", index=False)
Answer from betelgeuse on Stack ExchangeHere is a simple solution: Replace all \n with \\n before saving to CSV. This will preserve the newline characters.
df.loc[:, "Column_Name"] = df["Column_Name"].apply(lambda x: x.replace('\n', '\\n'))
df.to_csv("df.csv", index=False)
I assume that you want to keep the newlines in the strings for some reason after you have loaded the csv files from disk. Also that this is done again in Python. My solution will require Python 3, although the principle could be applied to Python 2.
The main trick
This is to replace the \n characters before writing with a weird character that otherwise wouldn't be included, then to swap that weird character back for \n after reading the file back from disk.
For my weird character, I will use the Icelandic thorn: Þ, but you can choose anything that should otherwise not appear in your text variables. Its name, as defined in the standardised Unicode specification is: LATIN SMALL LETTER THORN. You can use it in Python 3 a couple of ways:
weird_literal = 'þ'
weird_name = '\N{LATIN SMALL LETTER THORN}'
weird_char = '\xfe' # hex representation
weird_literal == weird_name == weird_char # True
That \N is pretty cool (and works in python 3.6 inside formatted strings too)... it basically allows you to pass the Name of a character, as per Unicode's specification.
An alternative character that may serve as a good standard is '\u2063' (INVISIBLE SEPARATOR).
Replacing \n
Now we use this weird character to replace '\n'. Here are the two ways that pop into my mind for achieving this:
using a list comprehension on your list of lists:
data:new_data = [[sample[0].replace('\n', weird_char) + weird_char, sample[1]] for sample in data]putting the data into a dataframe, and using replace on the whole
textcolumn in one godf1 = pd.DataFrame(data, columns=['text', 'category']) df1.text = df.text.str.replace('\n', weird_char)
The resulting dataframe looks like this, with newlines replaced:
text category
0 some text in one line 1
1 text withþnew line character 0
2 another newþline character 1
Writing the results to disk
Now we write either of those identical dataframes to disk. I set index=False as you said you don't want row numbers to be in the CSV:
FILE = '~/path/to/test_file.csv'
df.to_csv(FILE, index=False)
What does it look like on disk?
text,category
some text in one line,1
text withþnew line character,0
another newþline character,1
Getting the original data back from disk
Read the data back from file:
new_df = pd.read_csv(FILE)
And we can replace the Þ characters back to \n:
new_df.text = new_df.text.str.replace(weird_char, '\n')
And the final DataFrame:
new_df
text category
0 some text in one line 1
1 text with\nnew line character 0
2 another new\nline character 1
If you want things back into your list of lists, then you can do this:
original_lists = [[text, category] for index, text, category in old_df_again.itertuples()]
Which looks like this:
[['some text in one line', 1],
['text with\nnew line character', 0],
['another new\nline character', 1]]
csv writer: append doesn't go to new line
Add new line to output in csv file in python - Stack Overflow
Python, write a new line in a CSV-File - Stack Overflow
Python inserts newline by writing to csv - Data Science Stack Exchange
So, basically, I've got this code:
new_list_csv = []
def add_it():
#gets the values from Entry
#three different tk Entrys generate three different values
name= self.e.get()
ex= self.e_x.get()
ey= self.e_y.get()
#adds them to the new list
new_list_csv.append(name)
new_list_csv.append(ex)
new_list_csv.append(ey)
#appends new characters to csv file
with open ('chr_list_copy.txt', 'a', newline='') as write_obj:
csv_writer = csv.writer(write_obj)
csv_writer.writerow(new_list_csv)and the csv file looks like this:
name,x,y Ergo,1,1 Sum,5,5 Name12,1,2 Name34,3,4 Name56,5,6 Name78,7,8 Name910,9,10 Name1112,11,12 Name1314,13,14
if I try and add [Otto,6,9], I get:
name,x,y [...] Name1314,13,14Otto,6,9
instead of:
name,x,y [...] Name1314,13,14 Otto,6,9
I've used this very same structure for the rest of my code, but in this specific instance, it doesn't work.
Hard to answer your question without knowing the format of the variable census_links.
But presuming it is a list that contains multiple links composed of strings, you would want to parse through each link in the list and append a newline character to the end of a given link and then write that link + newline to the output file:
file = open('C:/Python34/census_links.csv', 'a')
# Simulating a list of links:
census_links = ['example.com', 'sample.org', 'xmpl.net']
for link in census_links:
file.write(link + '\n') # append a newline to each link
# as you process the links
file.close() # you will need to close the file to be able to
# ensure all the data is written.
E. Ducateme has already answered the question, but you could also use the csv module (most of the code is from here):
import csv
# This is assuming that “census_links” is a list
census_links = ["Example.com", "StackOverflow.com", "Google.com"]
file = open('C:\Python34\census_links.csv', 'a')
writer = csv.writer(file)
for link in census_links:
writer.writerow([link])
For a uni assignment, I need to process a csv file. However these csv files were written such that the EOL character in every case is a carriage return (\r) character. The assignment prompt says that I cannot read in the file using the csv module as it will not recognise the \r character and thus the requirement of the assignment is that, in my Python script, I change the carriage return to a newline or any other EOL character that Python's csv module can recognise as a valid EOL. How would I do this? Also why doesn't the csv module recognise the carriage return? (Sorry for poor English and/or bad formatting)
Python 3:
The official csv documentation recommends opening the file with newline='' on all platforms to disable universal newlines translation:
with open('output.csv', 'w', newline='', encoding='utf-8') as f:
writer = csv.writer(f)
...
The CSV writer terminates each line with the lineterminator of the dialect, which is '\r\n' for the default excel dialect on all platforms because that's what RFC 4180 recommends.
Python 2:
On Windows, always open your files in binary mode ("rb" or "wb"), before passing them to csv.reader or csv.writer.
Although the file is a text file, CSV is regarded a binary format by the libraries involved, with \r\n separating records. If that separator is written in text mode, the Python runtime replaces the \n with \r\n, hence the \r\r\n observed in the file.
See this previous answer.
While @john-machin gives a good answer, it's not always the best approach. For example, it doesn't work on Python 3 unless you encode all of your inputs to the CSV writer. Also, it doesn't address the issue if the script wants to use sys.stdout as the stream.
I suggest instead setting the 'lineterminator' attribute when creating the writer:
import csv
import sys
doc = csv.writer(sys.stdout, lineterminator='\n')
doc.writerow('abc')
doc.writerow(range(3))
That example will work on Python 2 and Python 3 and won't produce the unwanted newline characters. Note, however, that it may produce undesirable newlines (omitting the LF character on Unix operating systems).
In most cases, however, I believe that behavior is preferable and more natural than treating all CSV as a binary format. I provide this answer as an alternative for your consideration.
This problem occurs only with Python on Windows.
In Python v3, you need to add newline='' in the open call per:
Python 3.3 CSV.Writer writes extra blank rows
On Python v2, you need to open the file as binary with "b" in your open() call before passing to csv
Changing the line
with open('stocks2.csv','w') as f:
to:
with open('stocks2.csv','wb') as f:
will fix the problem
More info about the issue here:
CSV in Python adding an extra carriage return, on Windows
I came across this issue on windows for Python 3. I tried changing newline parameter while opening file and it worked properly with newline=''.
Add newline='' to open() method as follows:
with open('stocks2.csv','w', newline='') as f:
f_csv = csv.DictWriter(f, headers)
f_csv.writeheader()
f_csv.writerows(rows)
It will work as charm.
Hope it helps.
import csv
fh = open("employee.csv", "w", newline = '')
ewriter = csv.writer(fh)
empdata = [
['Empno', 'Name', 'Designation', 'Salary'],
[1001, 'Trupti', 'Manager', 56000],
[1002, 'Silviya', 'Clerk', 25000]
]
ewriter.writerows(empdata)
print("File successfully created")
fh.close()