pandoc

Author	SHA1	Message	Date
Jesse Rosenthal	9f6fd6139f	Docx reader: use all anchor spans for header ids. Previously we only used the first anchor span to affect header ids. This allows us to use all the anchor spans in a header, whether they're nested or not. Along with `62882f97`, this closes #3088.	2016-08-28 18:12:02 -04:00
Jesse Rosenthal	2893b0055a	Docx reader: Let headers use exisiting id. Previously we always generated an id for headers (since they wouldn't bring one from Docx). Now we let it use an existing one if possible. This should allow us to recurs through anchor spans.	2016-08-28 18:09:27 -04:00
Jesse Rosenthal	62882f976d	Docx reader: Handle anchor spans with content in headers. Previously, we would only be able to figure out internal links to a header in a docx if the anchor span was empty. We change that to read the inlines out of the first anchor span in a header. This still leaves another problem: what to do if there are multiple anchor spans in a header. That will be addressed in a future commit.	2016-08-28 18:03:09 -04:00
Jesse Rosenthal	9999db2e6c	StyleMap: export functions on StyleMap instances We're going to want `getMap` in the Docx Writer.	2016-08-15 12:04:20 -04:00
Jesse Rosenthal	347d716826	Docx parser: Use xml convenience functions The functions `isElem` and `elemName` (defined in Docx/Util.hs) make the code a lot cleaner than the original XML.Light functions, but they had been used inconsistently. This puts them in wherever applicable.	2016-08-13 13:46:21 -04:00
John MacFarlane	1955ee9c72	Merge pull request #3048 from tarleb/latex-mini-fix LaTeX reader: drop duplicate `*` in bibtexKeyChars	2016-08-11 21:15:27 +02:00
John MacFarlane	13424a2bd7	Merge pull request #3065 from tarleb/org-verse-indent Org reader: preserve indentation of verse lines	2016-08-09 21:33:24 +02:00
Albert Krewinkel	ba5b426ded	Org reader: ensure image sources are proper links Image sources as those in plain images, image links, or figures, must be proper URIs or relative file paths to be recognized as images. This restriction is now enforced for all image sources. This also fixes the reader's usage of uncleaned image sources, leading to `file:` prefixes not being deleted from figure images (e.g. `[[file:image.jpg]]` leading to a broken image `<img src="file:image.jpg"/>) Thanks to @bsag for noticing this bug.	2016-08-09 20:27:08 +02:00
Albert Krewinkel	13280a8112	Org reader: preserve indentation of verse lines Leading spaces in verse lines are converted to non-breaking spaces, so indentation is preserved. This fixes #3064.	2016-08-08 09:40:50 +02:00
John MacFarlane	0fbb676c81	MediaWiki reader: properly interpret XML tags in pre environments. They are meant to be interpreted as literal text in textile. Closes #3042.	2016-08-06 23:57:42 +02:00
John MacFarlane	124679fd63	Improved mediawiki reader's treatment of verbatim constructions. Previously these yielded strings of alternating Code and Space elements; we now incorporate the spaces into the Code. Emphasis etc. is still possible inside these. Closes #3055.	2016-08-06 23:41:03 +02:00
John MacFarlane	3a49439202	Fix for unquoted attribute values in mediawiki tables. Previously an unquoted attribute value in a table row could cause parsing problems. Fixes #3053 (well, proper rowspans and colspans aren't created, but that's a bigger limitation with the current Pandoc document model for tables).	2016-08-06 23:22:03 +02:00
Albert Krewinkel	f9afc0d378	LaTeX reader: drop duplicate `*` in bibtexKeyChars	2016-07-29 20:53:43 +02:00
John MacFarlane	27762affe3	Textile reader: disallow empty URL in explicit link. Closes #3036.	2016-07-22 15:45:03 -07:00
John MacFarlane	5f758970a5	Textile reader: support `bc..` extended code blocks. Also, remove trailing newline in code blocks (consistently with Markdown reader).	2016-07-22 15:32:50 -07:00
John MacFarlane	34533dd8d1	LaTeX reader: be more forgiving of non-standard characters. E.g. `^` outside of math. Some custom environments give these a meaning, so we should try not to fall over when we encounter them.	2016-07-20 11:36:50 -07:00
John MacFarlane	1b6c9733ee	LaTeX reader: more robust parsing of unknown environments. We no longer fail on things like `^` inside options for tikz. Closes #3026.	2016-07-20 11:18:24 -07:00
John MacFarlane	3263ed3c42	RST reader: use Div for admonitions. Previously blockquotes were used. Now a Div is used with class `admonition` and (if relevant) one of the following: `attention`, `caution`, `danger`, `error`, `hint`, `important`, `note`, `tip`, `warning`. `sidebar` is also put into a Div. Note: This will change rendering of RST documents! It should provide much more flexibility. Closes #3031.	2016-07-20 10:14:24 -07:00
John MacFarlane	e2d59461bb	Textile reader: improve definition list parsing. - Allow multiple terms (which we concatenate with linebreaks). - Fix exponential parsing bug (closes #3020 for real this time).	2016-07-19 09:03:15 -07:00
John MacFarlane	3490932d21	Textile reader: improved table parsing. We now handle cell and row attributes, mostly by skipping them. However, alignments are now handled properly. Since in pandoc alignment is per-column, not per-cell, we try to devine column alignments from cell alignments. Table captions are also now parsed, and textile indicators for thead and tfoot no longer cause parse failure. (However, a row designated as tfoot will just be a regular row in pandoc.)	2016-07-18 22:40:45 -07:00
John MacFarlane	d7396e73b4	Don't require haddock-library 1.4. Instead use CPP to work around version differences.	2016-07-15 12:04:00 -07:00
John MacFarlane	2f54de7cc4	Fixed compiler warnings.	2016-07-14 23:38:44 -07:00
John MacFarlane	c203ace130	Haddock reader - support math. The Haddock document model added elements for math in 1.4.	2016-07-14 23:38:20 -07:00
John MacFarlane	0b0a0e730f	Removed some redundant class constraints.	2016-07-14 08:54:06 -07:00
John MacFarlane	06a3e6a03f	Merge pull request #3019 from tarleb/org-verbatim-fix Org reader: fix parsing of verbatim inlines	2016-07-14 08:43:39 -07:00
John MacFarlane	00b11bcbcf	Fixed exponential parsing bug in textile reader. Closes #3020.	2016-07-14 08:42:38 -07:00
Albert Krewinkel	529146decf	Org reader: fix parsing of verbatim inlines Org rules for allowed characters before or after markup chars were not checked for verbatim text. This resultet in wrong parsing outcomes of if the verbatim text contained e.g. space enclosed markup characters as part of the text (`=is_substr = True=`). Forcing the parser to update the positions of allowed/forbidden markup border characters fixes this. This fixes #3016.	2016-07-14 13:33:25 +02:00
Albert Krewinkel	f417fecf5f	Org reader: replace ugly code with view pattern Some less-than-smart code required a pragma switching of overlapping pattern warnings in order to compile seamlessly. Using view patterns makes the code easier to read and also doesn't require overlapping pattern checks to be disabled.	2016-07-04 11:20:05 +02:00
John MacFarlane	e548b8df07	Merge pull request #3010 from tarleb/org-header-tree Org reader: support archived trees, headline levels export setting	2016-07-03 22:57:22 -07:00
John MacFarlane	4099b2dca4	Odt reader: Removed redundant Monoid constraints.	2016-07-03 22:47:32 -07:00
Albert Krewinkel	5ffa4abf72	Org reader: support headline levels export setting The depths of headlines can be modified using the `H` option. Deeper headlines will be converted to lists.	2016-07-03 23:28:45 +02:00
Albert Krewinkel	c1f6bd2640	Org reader: put export setting parser into module Export option parsing is distinct enough from general block parsing to justify putting it into a separate module.	2016-07-02 13:14:09 +02:00
John MacFarlane	e0cc9e4463	LaTeX reader: strip off double quotes around image source if present. Avoids interpreting these as part of the literal filename. See #2825.	2016-07-01 15:47:42 -07:00
Albert Krewinkel	c4cf6d237f	Org reader: support archived trees export options Handling of archived trees can be modified using the `arch` option. Archived trees are either dropped, exported completely, or collapsed to include just the header when the `arch` option is nil, non-nil, or `headline`, respectively.	2016-07-01 23:05:33 +02:00
Albert Krewinkel	1ebaf6de11	Org reader: refactor comment tree handling Comment trees were handled after parsing, as pattern matching on lists is easier than matching on sequences. The new method of reading documents as trees allows for more elegant subtree removal.	2016-07-01 23:05:32 +02:00
Albert Krewinkel	17484ed01a	Org reader: parse as headlines, convert to blocks Emacs org-mode is based on outline-mode, which treats documents as trees with headlines are nodes. The reader is refactored to parse into a similar tree structure. This simplifies transformations acting on document (sub-)trees.	2016-07-01 23:05:32 +02:00
Albert Krewinkel	2f8d6755f4	Org reader: improve tag and properties type safety Specific newtype definitions are used to replace stringly typing of tags and properties. Type safety is increased while readability is improved.	2016-07-01 23:05:32 +02:00
John MacFarlane	3429fa6438	LaTeX reader: fixed `\cite` so it is a NormalCitation not AuthorInText.	2016-06-29 07:59:00 -07:00
John MacFarlane	a349814665	Merge pull request #3001 from tarleb/org-figure-label Org reader: support figure labels	2016-06-26 17:51:51 -07:00
Albert Krewinkel	0f3f5ce1a1	Org reader: support figure labels Figure labels given as `#+LABEL: thelabel` are used as the ID of the respective image. This allows e.g. the LaTeX to add proper `\label` markup. This fixes half of #2496 and #2999.	2016-06-26 20:42:22 +02:00
John MacFarlane	38c97320ef	Textile reader: Fix overly aggressive interpretation as images. Spaces are not allowed in the image URL in textile. Closes #2998.	2016-06-25 14:04:47 -07:00
John MacFarlane	d283f9c864	Fixed RST links with no explicit link text. The link `<foo>`_ should have `foo` as both its link text and its URL. See RST spec at <http://docutils.sourceforge.net/docs/ref/rst/restructuredtext.html#embedded-uris-and-aliases> "The reference text may also be omitted, in which case the URI will be duplicated for use as the reference text. This is useful for relative URIs where the address or file name is also the desired reference text: See `<a_named_relative_link>`_ or `<an_anonymous_relative_link>`__ for details." Closes Debian #828167 -- reported by Christian Heller.	2016-06-25 10:56:37 -07:00
John MacFarlane	a820c1bd1c	Textile reader: fixed attributes. Attributes can't be followed by a space. So, _(class)emph_ but _(noclass) emph_ Closes #2984.	2016-06-23 10:28:54 -07:00
Jesse Rosenthal	032ba8dd0c	Docx reader: Add warning for advanced comment formatting. We can't guarantee we'll convert every comment correctly, though we'll do the best we can. This warns if the comment includes something other than Para or Plain.	2016-06-23 10:50:46 -04:00
Jesse Rosenthal	5f0cd89129	docx reader: enable warnings in top-level reader Previously we had only allowed for warnings in the parser. Now we allow for them in the `Docx.hs` as well. The warnings are simply concatenated.	2016-06-23 10:50:46 -04:00
Jesse Rosenthal	8bb739f7ff	Docx reader: add simple comment functionality. This adds simple track-changes comment parsing to the docx reader. It is turned on with `--track-changes=all`. All comments are converted to inlines, which can list some information. In the future a warning will be added for comments with formatting that seems like it will be excessively denatured. Note that comments can extend across blocks. For that reason there are two spans: `comment-start` and `comment-end`. `comment-start` will contain the comment. `comment-end` will always be empty. The two will be associated by a numeric id.	2016-06-23 10:50:46 -04:00
Albert Krewinkel	7df656089f	Org reader: remove partial functions Partial functions like `head` lead to avoidable errors and should be avoided. They are replaced with total functions. This fixes #2991.	2016-06-21 23:51:15 +02:00
Albert Krewinkel	29552eff3e	Org reader: support arbitrary raw inlines Org mode allows arbitrary raw inlines ("export snippets" in Emacs parlance) to be included as `@@format:raw foreign format text@@`. Support for this features is added to the Org reader.	2016-06-13 23:53:14 +02:00
Albert Krewinkel	8a9f5915ab	Org reader: add support for "Berkeley-style" cites A specification for an official Org-mode citation syntax was drafted by Richard Lawrence and enhanced with the help of others on the orgmode mailing list. Basic support for this citation style is added to the reader. This closes #1978.	2016-06-05 11:28:57 +02:00
Albert Krewinkel	06dfe3276d	Org reader: add semicolon to list of special chars Semicolons are used as special characters in citations syntax. This ensures the correct parsing of Pandoc-style citations: [prefix; @key; suffix] Previously, parsing would have failed unless there was a space or other special character as the last <prefix> character.	2016-06-05 11:28:57 +02:00

... 5 6 7 8 9 ...

1932 commits