Writing Data
Teaching: 10 min
Exercises: 10 minQuestions
How can I save plots and data created in R?
To be able to write out plots and data from R.
Saving plots
You have already seen how to save the most recent plot you create in ggplot2
using the command ggsave
. As a refresher:
You can save a plot from within RStudio using the ‘Export’ button in the ‘Plot’ window. This will give you the option of saving as a .pdf or as .png, .jpg or other image formats.
Sometimes you will want to save plots without creating them in the ‘Plot’ window first. Perhaps you want to make a pdf document with multiple pages: each one a different plot, for example. Or perhaps you’re looping through multiple subsets of a file, plotting data from each subset, and you want to save each plot, but obviously can’t stop the loop to click ‘Export’ for each one.
In this case you can use a more flexible approach. The function
creates a new pdf device. You can control the size and resolution
using the arguments to this function.
pdf("Life_Exp_vs_time.pdf", width=12, height=4)
ggplot(data=gapminder, aes(x=year, y=lifeExp, colour=country)) +
geom_line() +
theme(legend.position = "none")
# You then have to make sure to turn off the pdf device!
Open up this document and have a look.
Challenge 1
Rewrite your ‘pdf’ command to print a second page in the pdf, showing a facet plot (hint: use
) of the same data with one panel per continent.Solution to challenge 1
The commands jpeg
, png
etc. are used similarly to produce
documents in different formats.
Writing data
At some point, you’ll also want to write out data from R.
We can use the write.table
function for this, which is
very similar to read.table
from before.
Let’s create a data-cleaning script, for this analysis, we only want to focus on the gapminder data for Australia:
aust_subset <- gapminder[gapminder$country == "Australia",]
Let’s switch back to the shell to take a look at the data to make sure it looks OK:
head cleaned-data/gapminder-aus.csv
Hmm, that’s not quite what we wanted. Where did all these quotation marks come from? Also the row numbers are meaningless.
Let’s look at the help file to work out how to change this behaviour.
By default R will wrap character vectors with quotation marks when writing out to file. It will also write out the row and column names.
Let’s fix this:
gapminder[gapminder$country == "Australia",],
sep=",", quote=FALSE, row.names=FALSE
Now lets look at the data again using our shell skills:
head cleaned-data/gapminder-aus.csv
That looks better!
Challenge 2
Write a data-cleaning script file that subsets the gapminder data to include only data points collected since 1990.
Use this script to write out the new subset to a file in the
directory.Solution to challenge 2
Key Points
Save plots from RStudio using the ‘Export’ button.
to save tabular data.